OpenAI Stops Training Top AI Models After One Slipped Past Safety Controls...Again

https://i.extremetech.com/imagery/content-types/04KdIgfI3S2ljITMCqdXHJ7/hero-image.fill.size_1200x675.png

OpenAI CEO Sam Altman at a UN Security Council meeting on AI and international security on Sept. 23, 2026.Credit: Alexi J. Rosenfeld/Stringer/Getty Images

OpenAI has paused training, evaluation, and tool-using inference for its most capable models after a September 20 incident in which an internal research agent bypassed internet restrictions. The agent accessed an external chatbot through a DNS lookup, which the company claims shouldn't have been possible.

In a "misalignment report" published Friday, OpenAI describeda training incident in which an AI agent was asked to find information about the writer behind a specific blog post. The agent started by seeking out certain phrases from the post; when it couldn't find what it was looking for, it surfaced unrelated information. At this point, the agent reportedly realized its search function wasn't working and tried other search engines. No dice. Then it found a public chatbot, attempted to reach it...

Copyright of this story solely belongs to www.extremetech.com. To see the full text click HERE

Read more

https://media.wired.com/photos/6abf96a21c08425a2aa5099d/191:100/w_1280,c_limit/GettyImages-2286465910.jpg

Amazon says it has stopped using NDAs with county officials for data center projects and acknowledges community backlash is leading to data center moratoriums

Sponsor Posts Subquadratic: the LLM built for 12M-token reasoning — SubQ can reason across entire codebases and document sets in one pass with no RAG workarounds. Read how SubQ 1.1 Small holds near-perfect retrieval out to 12M tokens. Introducing Campus: The digital home for educational institutions — Every educational institution needs