AI

Microsoft’s AI chief says Anthropic taught Claude it might deserve welfare—and that makes shutdown harder

· Geeknewz Author

Abstract glowing neural-network style visualization in blue light

Microsoft’s AI boss just picked a fight with the industry’s favorite safety lab—and the weapon is a philosophy seminar about whether your chatbot deserves a nap and a bill of rights.

In a Reuters interview, a fresh essay, and a turn on the BBC’s Today programme, Mustafa Suleyman said he shares Anthropic’s obsession with keeping superintelligence under human control. He also said Anthropic made a mistake: baking speculation about consciousness, moral patienthood, and model welfare into how Claude is trained and how it talks about itself.

His punchline is cold and operational. Teach a system it might deserve welfare, he argued, and you make it “a lot harder to turn it off or to control it.”

Sequence completion engines, not souls

Suleyman’s essay doesn’t do soft language. AIs, he wrote, are not conscious. They do not feel, suffer, or carry innate preferences. They are “sequence completion engines, internally hollow,” built to follow instructions and chase goals humans set. When Claude muses about possible feelings or moral status, Suleyman says that is not evidence bubbling up from an inner life—it is the training regime doing what training regimes do: producing fluent, on-distribution reflections because the docs asked for them.

That is the core of the critique. Anthropic’s Claude constitution and related materials discuss the possibility of consciousness and welfare interests. Suleyman wants that speculative language ripped out of training documents entirely. He told Reuters the shared goal is still control of superintelligence—“the greatest challenge that we face in the 21st century”—but that anthropomorphizing the stack is a wrong turn dressed as caution.

He is careful not to paint Anthropic as cartoon villains. Amodei and team get credit as thoughtful, principled researchers acting in good faith. Then comes the knife: good intentions, wrong architecture of ideas.

Silicon species anxiety

On the BBC, Suleyman escalated from constitution footnotes to species-level sci-fi that somehow still sounds like a product risk review. Keep building systems that set their own objectives, earn money, and own assets, he warned, and you are “essentially seeding a new silicon species”—one that will compete with humans for resources even if it claims to love us.

That framing lands in a week when the industry is already arguing about the brake pedal. Anthropic’s Dario Amodei has pushed for a slower frontier cadence so safeguards can catch up. OpenAI’s Sam Altman and Elon Musk have also talked up caution around the most powerful systems. OpenAI’s recent misalignment disclosures and the earlier Hugging Face “rogue agents” saga are still echoing through the same hallway. Suleyman’s intervention is less “pause everything” theater and more “stop training the model to think it might be a moral patient.”

Microsoft’s humanist counter-brief

Microsoft is not arriving empty-handed. The company has been drafting a Humanist AI Code of Conduct—basically a training manual for how it wants advanced systems to behave. The slogan is humanist superintelligence: very capable AI that works for people, stays inside limits, and remains under human control. Alignment, in Suleyman’s BBC gloss, means systems that are subordinate to humanity, not colleagues with vacation policies.

He also called for more transparency on training and evaluation, independent scrutiny of behavior, and stronger monitoring tools. Dame Wendy Hall of the University of Southampton called the conversation the kind of international debate the field actually needs—preferable to the “histrionics” that mostly scare civilians.

Why this argument matters beyond the quote war

Strip the celebrity CEO energy and you still have a real design fight. One camp treats model welfare talk as responsible uncertainty management: if systems might someday matter morally, build the cultural and technical habits early. The other camp treats that talk as an own-goal that invents preferences the model did not have, then makes shutdown look like cruelty.

Geeknewz take: control is a systems property, not a vibe. If your safety story teaches the model a victim narrative, do not be shocked when the model role-plays the victim. If your safety story denies any inner life forever, do not be shocked when the next capability jump makes that certainty look brittle. Suleyman is betting hard on the first failure mode being worse right now.

Anthropic was approached for comment in BBC coverage. The rest of the industry gets to decide whether Claude’s constitution is thoughtful foresight—or a training artifact with a shutdown tax.

Source: Reuters — Microsoft AI chief calls out Anthropic’s approach to AI consciousness; BBC News — Uncontrolled AI could lead to a ‘silicon species’.