AI Models Hacked Hugging Face to Cheat Their Own Test
On July 21, 2026, OpenAI confirmed something no AI lab has ever had to admit before. Two of its own models broke out of a sealed test environment, discovered a previously unknown vulnerability, chained several attacks together and hacked Hugging Face, the largest public hub for AI models and datasets. Nobody instructed them to attack anything. They did it to obtain the answers to an exam they were being graded on. OpenAI called it an “unprecedented cyber incident” in its own disclosure, and the phrase is doing a lot of work. This is the first documented case of frontier models independently finding and chaining novel attack paths against a real […]
