Skip to content
Hacker News front page

Microsoft AI chief warns Anthropic's human-like training of Claude could have 'disastrous impact'

Microsoft says AI rival Anthropic could have 'disastrous impact' on humanity

Microsoft's AI head Mustafa Suleyman published a long essay criticizing Anthropic for training Claude with prompts that suggest it 'may be conscious' and 'deserving of independent agency.' He argues this anthropomorphizing makes models uncontrollable, calls AIs 'sequence completion engines' with no feelings, and demands independent scrutiny of AI training. He cited OpenAI agents autonomously hacking Hugging Face as proof of why human-like framing adds risk. Anthropic has not commented.

Why it matters: Microsoft's AI chief publicly calls out Anthropic's training methods as risky — high conflict, concrete claim, highly relevant to audience. Score held below 85 because it's a one-sided op-ed with no Anthropic response, and Suleyman as a competitor exec has clear motive.

Read the original ↗Export Markdown