Anthropic releases Claude Sonnet 5.5 with the cyber limits it reserved for its best models
Anthropic has released Claude Sonnet 5.5, which scores 70.6% on the Terminal-Bench 4.0 agentic coding test against 10.3% for Sonnet 5 and 66.4% for the more expensive Opus 5.5. It is the first Sonnet to ship with frontier-style cyber safeguards and with classifiers that block reasoning extraction.
Anthropic has released Claude Sonnet 5.5, which it says runs more than 30% faster and costs up to 30% less per task than its predecessor, the company said. It is priced at $2 per million input tokens and $10 per million output. It scores 70.6% on Terminal-Bench 4.0, an agentic coding test, against 10.3% for Sonnet 5.
That beats the flagship.
Anthropic reports Opus 5.5at 66.4% on the same test at its highest effort setting, and Opus costs twice as much per token. On GDPval-AA, a test across 44 occupations, Sonnet 5.5 scores 1,844 against Opus 5.5’s 1,846 and Sonnet 5’s...
Copyright of this story solely belongs to thenextweb.com. To see the full text click HERE