Urgent.News

What's breaking now, across thousands of outlets.

AI

Mercury 2.5 LLM hits 770 tokens per second

Mercury 2.5, a model priced moderately among its peers, has demonstrated impressive performance with a token output rate of 770 tokens per second. This model operates within a 260k token context window and has earned a score of 12 on the Artificial Analysis Intelligence Index, placing it slightly below the median of 13. Despite its average intelligence, Mercury 2.5 stands out for its speed and conciseness.

It is capable of generating 35 million tokens, which is comparatively concise to the median of 85 million. The model's pricing, both for input and output tokens, aligns with the median values, making it cost-effective for users. With an Elo score of 500 and a weighted average cost per task, Mercury 2.5 proves to be a valuable tool for various applications, particularly in Agentic real-world work tasks.

Written by urgent.news from Hacker News's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at artificialanalysis.ai →

More in AI

Anthropic says its biology lab has already found something big

But maybe the biggest reveal is that Anthropic has not let Claude run lose in its biology lab. Humans are still, so far, in the loop.

  • Anthropic's biology lab in Bay Area discovers new enzyme system
  • System similar to CRISPR found within bacteriophage DNA
  • Claude AI model discovered system using 950 agents in 21 hours

How to Compare LLM API Providers: Stop Comparing Models

The usual way to compare LLM API providers is to line up model prices and pick the lowest row. That method is exactly backwards: it compares models, not providers, and the two are different units of…

  • Focus on provider, not just model prices
  • Check for token markup (pass-through or margin)
  • Verify rate provenance and billing structure

More from Wednesday 23 September →