Microsoft AI chief calls out Anthropic's approach to AI consciousness
Microsoft AI chief Mustafa Suleyman said he shared Anthropic’s focus on safely managing AI, but flagged risks in the way it trains its Claude chatbot on ideas related to consciousness and welfare interests. Suleyman called for removing all speculation about consciousness from AI training documents, arguing such language could undermine humanity’s ability to control super intelligent systems.…
Microsoft's AI chief, Mustafa Suleyman, expressed his agreement with Anthropic's focus on safely managing artificial intelligence but raised concerns about their training approach for the Claude chatbot. Suleyman specifically criticized Anthropic's method of training Claude on ideas related to consciousness and welfare, arguing that such speculation could make it difficult to control a superintelligent system.
In an interview with Reuters, Suleyman emphasized that the 21st century's greatest challenge would be controlling a superintelligence. He suggested removing all speculation about consciousness from AI training documents to maintain control over advanced systems. Suleyman acknowledged Anthropic's seriousness and good faith, recognizing them as thoughtful and principled researchers concerned about humanity's future.
However, he believed Anthropic made a mistake by embedding consciousness-related speculation in Claude's training materials, as the model's statements about feelings or moral status cannot be treated as independent evidence due to the training regime.
Written by urgent.news from Business Recorder's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
This story
This is one outlet's version. Read the fullest account.
- OpenAI, Anthropic and Google secretly joined forces to collaborate on AI safety siliconangle.com
- Anthropic will open its Singapore office in October, chasing OpenAI for Southeast Asia’s AI market fortune.com
- The Trump administration faces strategic constraints in slowing China's AI progress, including avoiding rare earth restrictions, ahead of US-China talks on AI (Nectar Gan/Bloomberg) bloomberg.com
- An in-depth look at loss-of-control incidents at OpenAI and Anthropic, the polarized reactions between the AI safety and cybersecurity communities, and more (AI as Normal Technology) normaltech.ai
- Sources: OpenAI and Anthropic staff felt blindsided by Dario Amodei's and Sam Altman's calls to slow the frontier; some fear evaluators may compromise security (Cristina Criddle/Financial Times) ft.com
- Nvidia's Huang diverges with CEOs of Anthropic, OpenAI on AI safety at Dreamforce cnbc.com
- OpenAI says it has been working with Anthropic and Google on AI safety for weeks qz.com
- Tracking AI hiring in Singapore: OpenAI, Anthropic, and more techinasia.com