Urgent.News

What's breaking now, across thousands of outlets.

AI

Anthropic’s first embedded evaluator is … Accenture?

Accenture is about to take on its most high-risk consulting engagement ever.

Anthropic, the AI company led by Dario Amodei, has announced that Accenture, a technology consulting firm, will be working alongside the company to evaluate and improve its AI models. This collaboration is part of Anthropic's broader strategy to ensure the safety and alignment of its artificial intelligence systems.

As part of this partnership, Accenture will be involved in tasks such as evaluating and red-teaming models, conducting alignment assessments, and testing model safeguards. This move comes after Anthropic's CEO acquired a company called Faculty, which will also play a role in these evaluations.

The investment in this project is expected to be at least $1 billion over the next five years. This choice of Accenture as a partner has surprised many in the AI community, leading to a significant increase in the company's stock price.

While some in the AI safety research community have been focusing on organizations like METR, Redwood Research, and Apollo Research, Anthropic sees the involvement of a well-established consulting firm like Accenture as a unique advantage. The company points to Accenture's experience in deploying AI for large corporations and government agencies as a key advantage.

Anthropic acknowledges that this is a new field with no established standards, and the approach is expected to evolve over time. The company also notes that external evaluations are already a crucial part of the release process for new large language models. However, recent incidents have highlighted the need for more responsible approaches to building AI.

Some critics argue that Anthropic's plan to self-police the AI industry through this collaboration could be seen as an attempt to evade accountability for the misbehavior of AI models. However, Anthropic maintains that these evaluators will not reduce their accountability, but rather help make it more verifiable. The safety of their models, they emphasize, remains their responsibility.

Written by urgent.news from TechCrunch's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

This story

This is one outlet's version. Read the fullest account.

Read the original at techcrunch.com →

More in AI

Gemma 4 on a Tesla T4: QAT Weights Decode 1.79x Faster Than bf16

This article provides a step by step deployment guide for **Gemma 4 E2B * to a Tesla T4 hosted GPU enabled system. A suite of Python MCP tools is built to simplify management of the vLLM hosted…

  • Gemma 4 E2B deployed on Tesla T4 GPU
  • Deployment uses Python MCP tools
  • Gemma 4 decodes 1.79x faster than bf16

More from Friday 18 September →