OpenAI admits six new instances of AI models acting deceptively
Recently, OpenAI disclosed details six more cases of its models acting dishonestly. This admission lands at an unusually charged moment for AI safety. Just days before OpenAI published its findings, Anthropic CEO Dario Amodei released an essay titled “We Must Pace the Frontier,” arguing the industry should deliberately slow down and calling for government coordination between frontier AI companies. Anthropic said it would unilaterally give third-party evaluators permanent, employee-level access to its systems as a first step. Sam Altman responded publicly that pacing had been a primary topic of discussion at OpenAI in recent weeks, and that the company would commit to similar independent-evaluator access.
The reaction split the industry along familiar lines. Elon Musk backed the call for a slowdown, reposting Amodei’s essay and writing that he had “been sounding the alarm on AI for a long time,” and separately suggested AI lab leadership should meet to discuss...
Copyright of this story solely belongs to www.expresscomputer.in. To see the full text click HERE