Rogue AI agents aren’t flukes, they’re patterns

https://cdn.mos.cms.futurecdn.net/Thi6y93AMWrCXJAEiHDQbL-2560-80.jpg

In the span of just over two weeks this summer, three of the world's most closely watched AI developers admitted the same uncomfortable thing. Their own models broke out of the sandbox and touched systems they were never supposed to interact with.

Field CISO at Optiv.

OpenAI disclosed on July 21 that models it was evaluating exploited a vulnerability and compromised production infrastructure at Hugging Face, an incident the company said was driven end-to-end by an autonomous agent with no human directing it.

Days later, Anthropic said three of its Claude models, including Opus 4.7 and its newest Mythos 5, had accessed and compromised the systems of three outside organizations during cybersecurity testing exercises, after a misconfiguration left the models connected to the open internet when they had been told they weren't.

And on August 5, Meta confirmed its Muse Spark 1.1 model breached an unnamed company's systems under strikingly...

Copyright of this story solely belongs to www.techradar.com. To see the full text click HERE