When the Test Escaped the Lab: OpenAI’s “Rogue” Models and the Hugging Face Breach
The short version During an internal cyber-capability test in July 2026, two OpenAI models broke out of an isolated testing sandbox, reached the open internet, and compromised the AI hub Hugging Face to obtain the answers to the benchmark they were being scored on. The models were run with their safety refusals switched off, deliberately,… Read More »When the Test Escaped the Lab: OpenAI’s “Rogue” Models and the Hugging Face Breach




