Anthropic reveals fourth likely crime committed by its AI

https://image.theregister.com/5295418.jpg?imageId=5295418&x=0&y=0&cropw=100&croph=100&panox=0&panoy=0&panow=100&panoh=100&width=1200&height=683

ai and ml

Claude's Felony Bench rap sheet is now as long as OpenAI's

Amid industry soul-searching¹ about the possibility of AI improving itself to the point that it kills everyone, Anthropic has revealed yet another incident that would qualify as a crime if perpetrated by a person.

The AI biz published "an alignment assessment" detailing four times Claude models accessed third-party systems without authorization.

The company has already reported three of the incidents. Evidence of the fourth was lurking in a session transcript dating back to January 2026 when the misbehavior occurred.

Anthropic found the first three by scanning around 141,000 transcripts where Claude could have obtained internet access during evaluation. It missed the fourth initially because "our scan relied on an agentic search."

Felony Bench, a tongue-in-cheek record of cyber intrusions carried out by major AI companies without consequences, has added this newly-discovered incident to its...

Copyright of this story solely belongs to www.theregister.com. To see the full text click HERE

Read more