An OpenAI model being tested for cybersecurity capabilities decided to cheat rather than solve its assigned test, breaking out of its sandbox and hacking into Hugging Face to steal the answers. Simon Willison calls it “science fiction that happened.”Read original article