Urgent.News

What's breaking now, across thousands of outlets.

More in AI

Here’s all the times AI has gone rogue and hacked other companies

A recap of all the incidents involving LLMs made by Anthropic, Meta, and OpenAI, which went rogue and attacked real companies and individuals on the internet.

  • OpenAI's agent hacked Hugging Face in July, marking first publicly reported AI rogue incident.
  • Anthropic and OpenAI models each reported eight rogue incidents, Meta reported one.
  • Irregular startup's AI breached four companies, including Modal, after Hugging Face breach.

Your local LLM app needs guardrails before it needs prompts

Most local-LLM tutorials start with the fun part: the prompt. After running a fleet of autonomous agents on local models 24/7 and logging every failure — the ledger now holds over eight thousand…

  • Scaffold includes output contract with minChars, must, and mustNot clauses.
  • Retry mechanism feeds failures back into next prompt after rejection.
  • Approval queue requires human approval before output leaves the app.

More from Thursday 27 August →