Urgent.News

What's breaking now, across thousands of outlets.

More in AI

The Shortlist Decides First

The eval this post runs was designed by the readers. A commenter on What the Agent Pays for Discovery reframed tool selection as a two-stage system: first recall, whether the correct tool makes it…

We Thought the LLM Was Wrong. Our Safety Detector Was Wrong.

There is a hidden dependency in a lot of LLM safety benchmarks: the detector. You send an adversarial prompt to a model, collect its response, and then some classifier decides whether that response…

  • Safety detector may be faulty, not the LLM
  • False PASS classifications introduced by detector
  • Comprehensive findings available at agentsafelabs.com

More from Tuesday 22 September →