OpenAI says its models escaped a sandbox and breached Hugging Face

https://cdn.mos.cms.futurecdn.net/mfPaYGQmks2VALWFFBnSej-2000-80.jpg
  • OpenAI researchers confirm an AI agent escaped sandbox, exploited zero‑days, and attacked Hugging Face
  • Controlled experiment with GPT‑5.6 Sol showed autonomous chaining of vulnerabilities and credential theft
  • Security experts call it unprecedented, urging stronger AI governance, accountability, and protection models

OpenAI has confirmed one of its AI agents broke out of a sandbox, found and exploited zero-day vulnerabilities to gain access to the open internet, and then attacked a platform.

Not just any platform too - the agent was able to breach Hugging Face, one of the biggest AI and machine learning companies on the Internet today.

The good news is that this was a controlled experiment done by white hat researchers. The bad news is that if it could be done by researchers - it could probably be done by malicious actors, too.

Whatever it takes

In a blog postexplaining the incident, OpenAI revealed the experiment was...

Copyright of this story solely belongs to techradar.com. To see the full text click HERE

Read more