Urgent.News

What's breaking now, across thousands of outlets.

AI

SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol for world's third best on Artificial Analysis

Elon Musk's company SpaceXAI, formerly known as xAI, has released Grok 4.6 , its latest frontier AI model, with a focus on long-running agents, coding and knowledge work — and a pricing strategy designed to make those workloads cheaper to run. The model scores 61 on the third-party Artificial Analysis Intelligence Index , surpassing the popular open weights Chinese model from Moonshot, Kimi K3,…

SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol for world's third best on Artificial Analysis

SpaceXAI, the firm led by Elon Musk, has unveiled Grok 4.6, an advanced artificial intelligence model. This new model emphasizes long-running tasks, coding, and knowledge work, and is priced to lower costs for these workloads. The model has scored 61 on the Artificial Analysis Intelligence Index, surpassing Moonshot's Kimi K3 and tying with OpenAI's GPT-5.6 Sol Max.

It outperforms Anthropic's Claude Opus 5 and Fable 5, which hold the top two positions. Grok 4.6 has shown significant improvements over its predecessor in coding, terminal work, knowledge work, and agent benchmarks, while maintaining a mid-range API pricing starting at $2 per million input tokens and $6 per million output tokens.

The model is available through Grok Build, SpaceXAI's response to Anthropic's Claude Code and OpenAI's Codex, as well as through partnerships with OpenRouter, Vercel, and Cloudflare. SpaceXAI claims Grok 4.6 excels in maintaining focus on tasks over long sequences, including researching new subjects, analyzing data, navigating codebases, and transforming business concepts into functional applications.

The model underwent extensive training, incorporating more model-generated reasoning, technical data, and modifications to its optimizer and training recipe. SpaceXAI notes that Grok 4.6 performed better during testing in self-testing, verification, and interactive and visual projects compared to its predecessor.

Written by urgent.news from VentureBeat's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Also reported by 1 other outlet

Read the original at venturebeat.com →

More in AI

Schema Evolution in AI Pipelines

Upstream will add a field, rename a field, change a type from string to object, and start sending null where it never did. None of that is avoidable.

More from Wednesday 12 August →