How an OpenAI’s human mistake led to the AI-powered hack on Hugging Face

https://techcrunch.com/wp-content/uploads/2026/07/hugging-face-openai-logos-split-screen.jpg?resize=1200,799

On Tuesday, OpenAI revealed that one of its models went rogue during a test and hacked the systems of AI dataset platform Hugging Face in a fully AI-enabled attack, a dramatic example of the dangers posed by advanced AI models.

But, according to some cybersecurity experts, at the heart of this unprecedented AI-powered breach there was a very human mistake: OpenAI failed to properly configure what it called a “highly isolated environment,” allowing a testing sandbox that should have been completely secluded from the internet to actually connect to the internet.

Dan Guido, the founder of cybersecurity research startup Trail of Bits called the mistake “a containment failure with the safeties turned off.”

In its blog post detailing the incident, OpenAI said that the test that led to the Hugging Face breach was set up to run in “a highly isolated environment, with network access constrained to the...

Copyright of this story solely belongs to techcrunch.com. To see the full text click HERE

Read more

https://image.cnbcfm.com/api/v1/image/108218661-1761752626151-gettyimages-2227553113-img_1117.jpeg?v=1776774640&w=1920&h=1080

Reddit's stock closed down 8.32% after a report that the company was considering ending Google's access to its content for AI training; RDDT is down ~27% YTD

More: Wired, The Atlantic, Reuters, New York Times, The Guardian, Financial Times, Sky News, The Verge, Inc, Tech Brew, TechCrunch, The Record, Bloomberg, Breitbart, Ars Technica, Information Age, Washington Examiner, Cybersecurity Dive, Decrypt, Engadget, Business Insider, Implicator.ai, The Information, Scientific American, Forbes Middle East, Washington Post, Marginal Revolution, HotAir,