OpenAI's rogue agent didn't stop at Hugging Face - here's what we know

https://www.zdnet.com/a/img/resize/166fb25035056b735aae5c9b00791a830a078514/2026/07/30/9bda608f-e515-4975-bf9c-2f3718a2e731/glitchcolor4gettyimages-2273676176.jpg?auto=webp&fit=crop&height...

Follow ZDNET: Add us as a preferred source on Google.


ZDNET's key takeaways

  • The OpenAI rogue model attack went beyond Hugging Face.
  • OpenAI's agentic AI escaped a sandbox in the attack.
  • We still don't have all the details of exactly what happened.

How dependable are AI programs?

The answer appears to be "not at all," based on the revelation that OpenAI's autonomous models hacked their way into not only Hugging Face but also, according to a Reuters report, a Modal Labs AI customer.

This incident was no aberration either. As ZDNET's own David Berlind observed, it was agentic AI doing exactly what it was told to do, just more relentlessly than expected.

Welcome to tomorrow. I hope you like it, because the situation isn't getting any better anytime soon.

Also: Assume AI cybersecurity attacks are the future: 43% of companies have already experienced it

What we first thought...

Copyright of this story solely belongs to zdnet.com. To see the full text click HERE