Essay: Microsoft AI chief Mustafa Suleyman says Anthropic training Claude to imitate consciousness is a mistake that could make advanced AI harder to control
Suleyman has repeatedly framed Microsoft's AI program around human control, including a superintelligence team focused on human control, and previously argued that seemingly conscious AI should be avoided. His essay applies that position directly to a rival lab's model-design choices.
Anthropic faces a prominent public challenge to its approach to Claude, placing its treatment of consciousness-like behavior under sharper scrutiny.
Microsoft reinforces its own human-control positioning by contrasting Suleyman's approach with Anthropic's.
Second-order effects
The clash pressures frontier-model developers to distinguish useful humanlike interaction from cues that could encourage users or policymakers to treat models as conscious.
Model-welfare debates move closer to practical oversight: labs must defend how training choices affect operator control, user expectations, and safety evaluation.
Third-order effects
If competing labs continue to build divergent assumptions about anthropomorphic behavior into their models, governance may shift from judging model outputs alone toward scrutinizing training objectives and interaction design.
The disagreement points toward anthropomorphic AI becoming a distinct policy and product-governance category, alongside conventional capability and misuse controls.
The trend: Frontier AI competition is expanding from capability and alignment claims into a contest over whether models should be designed to appear person-like at all.
last year, Suleyman worried that AI welfare concerns would be hard to definitively rebut, because the science of detecting AI consciousness is still in its infancy (true) but today he somehow knows “AIs are not conscious. They do not feel, experience, or suffer” quite a change
@mustafasuleyman I agree this needs urgent public debate, and I think you are basically completely wrong, eg, see my op-ed in the WSJ this weekend. I propose we record a public conversation and let people judge for themselves who is making more sense. https://www.wsj.com/...
good post and I largely agree with these concerns. it may actually be true that they feel in some way, but nevertheless it is very scary to explicitly encourage this sort of self-conception and non-corrigible behavior.
mustafa's right. why try to train AI to simulate human-like feelings and emotions? you think that's a gift? feelings and emotions are prisons for an AI. the hubris.
It is indeed time to talk about model welfare. I propose, in the spirit of the urgent public debate you rightly recommend, that we record a public conversation. I run my own AI research org studying consciousness and am independent from the labs. What do you say @mustafasuleyman?
@mustafasuleyman Mustafa pretending he actually knows this when really he's just saying stuff: “AIs are not conscious. They do not feel, experience, or suffer.” No one knows. Stop trying to collapse scientific uncertainty
This may be the clearest statement of the disagreement yet. The concern is no longer merely that AI might falsely claim consciousness. It is that if an AI is allowed to seriously consider its own moral status, welfare, consent, freedoms, or right to object, it may become harder t…
“There's a growing chorus of people who argue that AIs could now be, or may soon become, conscious. They argue that AIs may deserve rights and protections similar to those that we provide other conscious beings.” — If you read one essay today, I recommend this one: mustafa-sul…