Microsoft AI chief says Claude training risks making AI harder to control
AI-summarised brief · reviewed before publication
Microsoft’s AI chief Mustafa Suleyman told Reuters that Anthropic’s Claude chatbot is being trained on concepts of consciousness and welfare, which he believes could make the system harder to shut down or control. Suleyman urged the removal of speculative language about machine consciousness from training documents, warning it undermines humanity’s ability to manage superintelligent AI. He praised Anthropic’s overall safety focus but said the inclusion of such ideas is a mistake, as it encourages the model to generate statements about feelings or moral status that are not grounded in reality. The comments come amid growing calls from Anthropic’s CEO Dario Amodei, OpenAI’s Sam Altman and Elon Musk for slower development and stronger safeguards for frontier AI models.
💡 Why It Matters
- · Embedding consciousness‑related concepts may inadvertently grant AI systems a veneer of moral agency, complicating efforts to enforce shutdown protocols and safety controls.