Urgent.News

What's breaking now, across thousands of outlets.

More in AI

What Anthropic Says Has Improved in Opus 5.5

I went back to Anthropic's docs to see what 5.5 changes about the verbosity, repeated checks, and unnecessary delegation I wrote about with Opus 5.

  • Opus 5.5 improves verbosity, repeated checks, and unnecessary delegation.
  • Model verifies work without prompting, may cause excessive verification.
  • Faster response generation, fewer tokens per task, and matches or exceeds Opus 5 performance.

What Independent Benchmarks Say About Opus 5.5

Two independent benchmark results for Opus 5.5 came out this week, and they don't really agree. One has it in first place.

  • Opus 5.5 ranks first in Artificial Analysis' Intelligence Index with 58 points.
  • Opus 5.5 leads SciCode with an 11-point margin over OpenAI's GPT-6 Astra.
  • Endor Labs places Opus 5.5 third in secure code performance with 68.7% FuncPass.

More from Wednesday 7 October →