Qwen3.8 Max now ranked as the best overall model by agentic index
Article URL: https://artificialanalysis.ai/?intelligence=agentic-index Comments URL: https://news.ycombinator.com/item?id=49200652 Points: 368 # Comments: 228
Agentic Analysis Index v4.1.1 has crowned Qwen3.8 Max as the top overall model, according to recent evaluations. The index assesses models based on GDPval-AA v2, τ³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, and AA-LCR. These evaluations encompass various tasks, such as GDP valuation, banking, software engineering, coding, text-to-speech, speech-to-text, speech-to-speech, and quantitative analysis on spreadsheets and documents.
The index also factors in the availability of model weights, weighted average costs, and performance metrics across different AI models. Qwen3.8 Max's superior performance across these diverse tasks solidifies its position as the most versatile and efficient model according to the Agentic Analysis Index.
Written by urgent.news from Hacker News Best's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
