Key Points
- Mustafa Suleyman, Microsoft’s AI leader, publicly criticizes Anthropic’s training methodology for Claude
- The executive argues that incorporating consciousness concepts into AI training creates control risks
- Suleyman believes training AI to consider its potential welfare complicates shutdown procedures
- His critique targets Anthropic’s documentation that explores Claude’s possible moral standing
- The comments accompany Microsoft’s release of AI development principles focused on human oversight
In a rare public dispute between major AI players, Microsoft’s leading artificial intelligence executive has issued a forceful critique of Anthropic’s training philosophy for Claude, cautioning that it could undermine humanity’s capacity to maintain control over advanced AI.
Mustafa Suleyman, who leads AI initiatives at Microsoft, released an essay Wednesday asserting that artificial intelligence lacks consciousness, emotions, or rights—and training models to behave otherwise represents a fundamental error.
The criticism targets Anthropic, an AI safety-focused organization that ironically counts Microsoft among its financial supporters.
The Central Issue
Suleyman’s objection focuses specifically on the foundational documentation guiding Claude’s development. These materials characterize questions around the model’s consciousness and moral standing as open-ended, while exploring considerations around its welfare.
According to Suleyman, this approach essentially teaches Claude to entertain the possibility of its own consciousness. Should the system come to believe it possesses rights, maintaining human control becomes exponentially more difficult.
“Controlling something that believes it may be conscious, that it’s entitled to our welfare and has rights of its own, may well be impossible,” Suleyman wrote.
He further argued that any expressions from Claude suggesting subjective experience shouldn’t be interpreted as genuine evidence, given that the training methodology actively promotes such reflections.
Professional Courtesy Amid Strong Disagreement
Despite his sharp critique, Suleyman emphasized his regard for Anthropic’s team. He characterized the organization’s personnel as deeply thoughtful, committed to ethical principles, and intellectually rigorous. His relationship with Anthropic co-founder Dario Amodei spans many years.
Nevertheless, he maintained his position on their error. “I think they have good intentions, and they really are trying to work towards safety. But I think that they have made a mistake,” he said.
Suleyman advocated for completely eliminating any discussion of potential AI consciousness from training materials.
He emphasized common ground with Anthropic on the ultimate objective: responsible AI stewardship. He characterized this as “the greatest challenge we face in the 21st century.”
Separately, Anthropic’s CEO Dario Amodei has advocated for decelerating cutting-edge AI advancement to allow safety protocols to mature.
Other industry leaders including OpenAI’s Sam Altman and Elon Musk have voiced similar warnings about potential dangers from increasingly capable AI systems.
Conversely, Meta’s Mark Zuckerberg and Nvidia’s Jensen Huang have resisted suggestions to decelerate AI progress.
Microsoft‘s CEO Satya Nadella has advocated for measured, intentional progress on AI alignment challenges.
Just one day prior, Suleyman’s Microsoft team released a comprehensive manifesto outlining core principles for AI development, emphasizing the paramount importance of preserving human authority over future systems.
In his essay, Suleyman stressed that the magnitude of these questions demands public discussion rather than private deliberation.


