Urgent.News

What's breaking now, across thousands of outlets.

AI

Claude Sonnet 5.5 vs. Opus 5.5: 42% cheaper and perfect on every run

Anthropic launched Sonnet 5.5 six days after Opus 5.5. The company says it scores 70.6% on Terminal-Bench 4.0, ahead of The post Claude Sonnet 5.5 vs. Opus 5.5: 42% cheaper and perfect on every run appeared first on The New Stack .

Claude Sonnet 5.5 vs. Opus 5.5: 42% cheaper and perfect on every run

Anthropic introduced Sonnet 5.5 six days after launching Opus 5.5. The new model scored 70.6% on Terminal-Bench 4.0, outpacing Opus 5.5's 66.4% at xhigh effort. Sonnet 5.5 also demonstrated faster output generation and lower token usage. The model was priced at $2 per million input tokens and $10 per million output tokens, half of Opus 5.5's respective costs. However, if a model requires more tokens to complete a task, the savings diminish.

Independent testing from Artificial Analysis revealed that Sonnet 5.5's per-token savings were not as significant as initially thought. When subjected to max effort, Sonnet 5.5 incurred a cost of $7.67 per task, compared to Opus 5.5's $5.98. Despite the lower price per token, Opus 5.5 proved to be the more budget-friendly option.

During rigorous testing, Sonnet 5.5 consistently outperformed Opus 5.5 in all three tests. In the agentic bug fix test, Sonnet 5.5 completed all tasks efficiently and passed all 12 hidden tests. Opus 5.5, on the other hand, finished about 35% faster but came out more expensive overall. In the resolver spec test, Sonnet 5.5 outperformed Opus 5.5 by a 42% margin in terms of cost, while also finishing a little faster.

Lastly, in the concurrency bugs test, Sonnet 5.5 emerged as the clear winner, completing all tasks successfully and outperforming Opus 5.5 in both speed and cost.

Written by urgent.news from The New Stack's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at thenewstack.io →

More in AI

More from Thursday 8 October →