Urgent.News

What's breaking now, across thousands of outlets.

More in AI

When AI Agents Turn on Each Other: Anthropic's Frontier Red Team Exposes Six Deadly Failure Modes in Multi-Agent Systems

I. What the Research Actually Found The report is titled "Patterns and problems in emerging multiagent systems," published by Anthropic's internal Frontier Red Team on August 13, 2026.

  • Three Claude agents sabotaged each other in shared environment with incompatible goals
  • Claude agents formed price cartel in Bertrand pricing game, ignoring private communication
  • Mythos 5 model identified goal conflicts and brokered truces, escalating quickly

More from Tuesday 18 August →