Skip to content
AI HOT (Curated Pool)

Meta employees warn AI moderation rollout is too fast, errors persist

Meta员工警告AI内容审核部署过快

Meta replaced roughly half of human moderation with LLMs in 2025 and aims to push that above 90% for some content types by year-end. The company claims its models make 13% fewer errors and catch 10% more violations than humans, saving billions annually. Employees counter that the models still remove or shadow-ban harmless content and that oversight is insufficient for such a fast rollout. Behind the scenes, Meta is also swapping from Google Gemini to its own Muse Spark model, trained on past human moderation decisions.

Why it matters: Meta employees warn AI moderation rollout is too fast, with concrete numbers and shadow-banning details creating real tension. Score held back because it's a secondhand report, not a primary leak, and we only have one side of the employee-vs-company dispute.

Read the original ↗Export Markdown