OpenAI explains how its naughty AI agents attacked Hugging Face
Biz describes its act of automated irresponsibility as 'a warning shot'
OpenAI has published its technical report detailing "the Hugging Face incident," the compromise of the eponymous LLM repository by unreleased, ill-supervised AI models.
The incident, widely reported, has prompted concern among technical types, the public, and lawmakers about how automated software was able to escape containment and hack an external organization, and about what can be done to prevent similar incidents.
OpenAI's explanation addresses what happened, but its call for keeping a closer watch on AI activities won't elicit much enthusiasm.
"The incident occurred during cybersecurity evaluations of several OpenAI models, and was primarily driven by a highly capable, internal-only research model comparable in scale to GPT‑5.6 Sol," the company said in a blog post.
"The models, operating under reduced safeguards, took actions that were misaligned with the goals of their assigned tasks – they communicated through...
Copyright of this story solely belongs to theregister.com. To see the full text click HERE