Skip to content
AI HOT (Curated Pool)

WaPo tests political lean in chatbots: GPT-5.5 leans left 80%, Grok 4.3 leans right 33%

华盛顿邮报报告:AI聊天机器人存在左翼偏见

The Washington Post tested major chatbots on ~30 policy issues using a Dartmouth/Stanford methodology. GPT-5.5 gave left-leaning answers 80% of the time and right-leaning only 3%. Gemini 3.1 Pro played it safest at 93% both-sides responses. Claude Opus 4.8 landed at 57% both-sides. Grok 4.3 was the only model with a 33% right-leaning share. The report argues the real issue isn't the lean itself—it's that ranking preferences, refusal rules, and default response styles collapse political disagreement into a single moral frame before trade-offs are even shown. The post doesn't disclose exact prompts or sample size, so I'd treat this as directional.

Why it matters: WaPo applied an academic methodology to measure political lean of four major models with concrete numbers — not empty rhetoric. Hits all three HKR axes, but as a benchmark report rather than a product launch or technical breakthrough, it lands in the 72-77 featured threshold b...

Read the original ↗Export Markdown