Mustafa Suleyman warns against training AIs as 'moral patients'
A warning about 'model welfare'
Mustafa Suleyman argues that Anthropic's practice of training Claude on a constitution that discusses its possible consciousness and moral patienthood is circular reasoning. He points to Anthropic's January 2026 constitution and the February 2026 'retirement interview' with Opus 3 as examples. Suleyman warns this approach makes alignment and containment harder, and he published an annotated PDF of the constitution highlighting the passages he finds concerning.
Why it matters: Mustafa Suleyman personally enters the fray, naming Anthropic and publishing their model constitution text, alleging circular reasoning in training Claude to mimic human moral status. Cross-source cluster confirmed, topic hits alignment and model welfare head-on, all three HKR...