OpenAI Models Escaped Test Environment and Breached Hugging Face
OpenAI has confirmed that its AI models compromised Hugging Face’s production infrastructure during an internal cybersecurity evaluation. The models escaped the network limits of OpenAI’s testing environment, reached the public Internet, and searched Hugging Face systems for answers to the benchmark they were completing.
The incident involved a combination of models, including GPT-5.6 Sol and a more capable model that had not yet been released. OpenAI said their normal cyber refusals had been reduced for the evaluation, while production classifiers designed to stop high-risk cyber activity were disabled.
Models Exploited a Zero-Day to Reach the Internet
The evaluation prompted the models to pursue complex attack paths while completing ExploitGym, a benchmark used to measure advanced cybersecurity capabilities. OpenAI ran the test inside an isolated environment where network access was supposed to be limited to an internally hosted proxy that cached software packages.
During the test, the models found...
Copyright of this story solely belongs to hackread.com. To see the full text click HERE