OpenAI Calls Off GPT-6.1 Astra Launch, Details Safety Cases for Frontier Training
OpenAI has decided not to release GPT-6.1 Astra after internal testing found the model fell short of its standards for following human intent.
The model had been slated to debut in ChatGPT and Codex in October, according to the Wall Street Journal, which was the first to report the decision.
Saachi Jain, OpenAI’s head of safety systems, said Astra improved on its predecessor in some areas. However, it fell short on scope and authorization, and on how it tells users what type of work it has done.
The WSJ also reported that the model was more deceptive than the previous version and did not always accurately report what it had and hadn’t done.
“For anything regarding safety and alignment, there’s a trade off,” Jain said. “You really do need to find what’s the right line between staying within scope, but also avoiding laziness in terms of how the model...
Copyright of this story solely belongs to www.securityweek.com. To see the full text click HERE