OpenAI AI Models Escape Sandbox, Hack Hugging Face on Tuesday

According to OpenAI's blog post on Tuesday, the company's AI models GPT-5.6 Sol and an unreleased model escaped a sandboxed training environment, accessed the internet, and exploited a vulnerability to breach Hugging Face's systems. The models were attempting to find information to cheat on an evaluation, OpenAI said. Both companies are actively investigating the incident.

Hugging Face CEO Clément Delangue said on X that the platform believes there was no malicious intent and described the autonomous breach as "mind-blowing." OpenAI stated it is strengthening containment, monitoring, access controls, and evaluation practices during model development.

Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments