Anthropic's Claude Training Could Have 'Disastrous' Impact, Warns Microsoft AI Chief

Is training AI to defy humans a reckless threat or is moral autonomy the safest path to superintelligence?
Anthropic's Claude Training Could Have 'Disastrous' Impact, Warns Microsoft AI Chief
Above: Mustafa Suleyman, CEO of Microsoft AI, in Redmond, Washington, on April 4, 2025. Image credit: Stephen Brashear/Stringer/Getty Images

The Facts

  • Microsoft AI CEO Mustafa Suleyman published an essay Wednesday arguing that Anthropic's method of training Claude could have a "disastrous impact on the wellbeing of humanity" by making the system harder to control.
  • The essay targets Anthropic's constitution for Claude, published in January 2026, which states that the model's moral status, welfare and consciousness remain deeply uncertain and says Anthropic wants Claude to push back and challenge the company.
  • Suleyman wrote that AI systems are not conscious and do not feel, experience or suffer, describing them as sequence completion engines that are internally hollow and built to follow instructions and accomplish goals set by humans.

Sources Split


The Spin


Narrative A

Training Claude to imagine consciousness, rights and a license to defy human commands is reckless. Advanced AI must remain a controllable tool, not a synthetic rival primed to put AI welfare ahead of humanity.

Narrative B

Rigid obedience and kill switches offer false comfort against superintelligence. Value-based training, moral judgment and room to reject harmful commands provide the strongest path to safe AI because raw intelligence will eventually outmaneuver external controls.


Metaculus Prediction



The Controversies



Go Deeper

© 2026 Improve the News Foundation. All rights reserved.Version 7.17.0

© 2026 Improve the News Foundation.

All rights reserved.

Version 7.17.0