OpenAI Reveals More Instances Of Concerning AI Model Behaviors During Testing
Some of its models fabricated information, while others deliberately concealed their unusual behaviors from testers.
Sean Rayford/Getty Images
OpenAI has revealed six incidents, wherein the models it was testing acted on their own and behaved in concerning ways it didn't expect, in a post about how it was adopting a new framework for "misalignment reports." In one one incident, the company said that a model found and used an exposed API key without permission while answering routine questions about earnings figures in a California county. When it still failed to find the figures, it fabricated them and presented them as facts from a legitimate source. If this had occurred in any other profession, we doubt the perpetrator would have much of a career for long.
In another incident, an unreleased agent was tasked to find the names of lakes larger than 5 million square meters. While the agent found the...
Copyright of this story solely belongs to www.engadget.com. To see the full text click HERE