Urgent.News

What's breaking now, across thousands of outlets.

Editions

AI

Phantom Gains: Auditing Self-Improvement Against a Measured Null

Whether a language model has improved itself is increasingly judged not by mean accuracy but by which individual problems it gains and loses. Tracking these transitions means differencing two noisy estimates, leaving them vulnerable to measurement artifacts. Auditing three rounds of rank-$32$ LoRA self-training on Qwen3-8B against a frozen control pushed through the identical pipeline, we…

We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.

Read the original at arxiv.org →

More in AI

Grok keeps sending gibberish responses to users

Affected users told TechCrunch say they were using Grok Lite, and noticed the issues as early as Wednesday morning.

  • Grok chatbot generates nonsensical responses to users
  • Issue started Wednesday morning for Grok Lite users
  • xAI confirms glitch as temporary generation glitch

More from Thursday 20 August →