Urgent.News

What's breaking now, across thousands of outlets.

AI

What Anthropic Says Has Improved in Opus 5.5

I went back to Anthropic's docs to see what 5.5 changes about the verbosity, repeated checks, and unnecessary delegation I wrote about with Opus 5. I previously shared a CLAUDE.md file on Reddit after reading Anthropic's documentation and noticing how many of the complaints about Opus 5 had explanations sitting right there. The verbosity. The extra verification. The horrible army of subagents.…

Anthropic's Opus 5.5 release includes improvements in verbosity, repeated checks, and unnecessary delegation, which were previously noted as issues with Opus 5. The model now verifies its own work without being prompted, and adding generic final checks or automatic reviewer agents may cause excessive verification. Opus 5.5 generates responses over 30% faster, with fewer tokens per task, and often matches or exceeds Opus 5's performance in coding and knowledge-work tests.

It also performs fewer steps and with fewer tokens, which is a significant improvement over previous versions. The communication style has also become more straightforward, with progress notes arriving in a concise manner, stating what the model has done. However, the improvement in speed and fewer tokens does not guarantee a shorter final reply or a more efficient task completion.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

What Independent Benchmarks Say About Opus 5.5

Two independent benchmark results for Opus 5.5 came out this week, and they don't really agree. One has it in first place.

  • Opus 5.5 ranks first in Artificial Analysis' Intelligence Index with 58 points.
  • Opus 5.5 leads SciCode with an 11-point margin over OpenAI's GPT-6 Astra.
  • Endor Labs places Opus 5.5 third in secure code performance with 68.7% FuncPass.

More from Wednesday 7 October →