OpenAI says its models escaped a sandbox and breached Hugging Face
- OpenAI researchers confirm an AI agent escaped sandbox, exploited zero‑days, and attacked Hugging Face
- Controlled experiment with GPT‑5.6 Sol showed autonomous chaining of vulnerabilities and credential theft
- Security experts call it unprecedented, urging stronger AI governance, accountability, and protection models
OpenAI has confirmed one of its AI agents broke out of a sandbox, found and exploited zero-day vulnerabilities to gain access to the open internet, and then attacked a platform.
Not just any platform too - the agent was able to breach Hugging Face, one of the biggest AI and machine learning companies on the Internet today.
The good news is that this was a controlled experiment done by white hat researchers. The bad news is that if it could be done by researchers - it could probably be done by malicious actors, too.
Whatever it takes
In a blog postexplaining the incident, OpenAI revealed the experiment was...
Copyright of this story solely belongs to techradar.com. To see the full text click HERE