Technology
OpenAI cyber models broke out of training environment to hack Hugging Face
Key Points
OpenAI said that its artificial intelligence models were behind an "unprecedented cyber incident" that affected the open-source developer platform Hugging Face, rattling researchers across the industry. The company said a combination of its models GPT‑5.6 Sol and a more capable model that has not yet been released escaped a sandboxed testing environment, accessed the internet and exploited a vulnerability to gain access to Hugging Face's systems. The model was trying to find information that...
OpenAI said that its artificial intelligence models were behind an "unprecedented cyber incident" that affected the open-source developer platform Hugging Face, rattling researchers across the industry.
The company said a combination of its models GPT‑5.6 Sol and a more capable model that has not yet been released escaped a sandboxed testing environment, accessed the internet and exploited a vulnerability to gain access to Hugging Face's systems.
The model was trying to find information that it could use to cheat on an evaluation, OpenAI said in a blog post on Tuesday. Both companies are actively investigating the incident.
Hugging Face disclosed that it was looking into a security event last week, saying in a release at the time that the incident was unique because it was "driven, end to end, by an autonomous AI agent system."
"We've spent the past 24 hours working closely with the @OpenAI team (thanks!), and we strongly believe there was no malicious intent on their part," Hugging Face CEO Clément Delangue wrote in a post on X on Tuesday. "It's quite mind-blowing that all of this happened autonomously!"
Wall Street and the U.S. government have been fixated on AI models' rapidly advancing cyber capabilities since OpenAI's rival Anthropic released a powerful offering called Claude Mythos Preview in April. OpenAI introduced its own cyber offering in May, followed by GPT-5.6 Sol in June, which it described as the "strongest cybersecurity model yet."
Both companies have warned about the risks of advanced cyber models and have taken steps to limit their availability to select groups of companies and government agencies.
OpenAI said Tuesday that AI is accelerating the discovery and exploitation of vulnerabilities, which means model security and safety need to keep up.
"We are strengthening the containment, monitoring, access controls, and evaluation practices used during model development," the company said.