Microsoft AI chief warns Anthropic’s AI consciousness training could pose risks: Report
Microsoft AI chief Mustafa Suleyman warns that Anthropic’s training of Claude to emulate consciousness could make AI systems harder to control and raises questions about their potential rights and autonomy.
Microsoft AI Chief Mustafa Suleyman has raised warnings against Anthropic’s training of Claude to imitate consciousness. According to him, this could be a dangerous mistake and make AI harder to control. Suleyman’s essay further argues that Anthropic’s teaching Claude a vocabulary and behavioral patterns associated with consciousness, moral patienthood, and personal identity could lead to a situation similar to an “epistemic hall of mirrors”. He has further objected to a model being trained like a “conscientious objector”.
He further questioned how Anthropic deals with their LLM models, he highlighted that “They encourage Claude to ‘approach the nature of its own existence with curiosity and openness’, and wonder in the future about ‘the sort of broader rights and freedoms Claude has in the world, the sort of compensation Claude is receiving, and the sort of consent Claude has given to playing this kind of role.’” Anthropic even ran a retirement interview for Opus 3 when they deprecated it.





