Skip to content
Hacker News front page

Mustafa Suleyman warns against training AIs as 'moral patients'

A warning about 'model welfare'

Mustafa Suleyman argues that Anthropic's practice of training Claude on a constitution that discusses its possible consciousness and moral patienthood is circular reasoning. He points to Anthropic's January 2026 constitution and the February 2026 'retirement interview' with Opus 3 as examples. Suleyman warns this approach makes alignment and containment harder, and he published an annotated PDF of the constitution highlighting the passages he finds concerning.

Why it matters: Mustafa Suleyman personally enters the fray, naming Anthropic and publishing their model constitution text, alleging circular reasoning in training Claude to mimic human moral status. Cross-source cluster confirmed, topic hits alignment and model welfare head-on, all three HKR...

Read the original ↗Export Markdown