Meta employees warn AI moderation rollout is too fast, errors persist
Meta员工警告AI内容审核部署过快
Meta replaced roughly half of human moderation with LLMs in 2025 and aims to push that above 90% for some content types by year-end. The company claims its models make 13% fewer errors and catch 10% more violations than humans, saving billions annually. Employees counter that the models still remove or shadow-ban harmless content and that oversight is insufficient for such a fast rollout. Behind the scenes, Meta is also swapping from Google Gemini to its own Muse Spark model, trained on past human moderation decisions.
Why it matters: Meta employees warn AI moderation rollout is too fast, with concrete numbers and shadow-banning details creating real tension. Score held back because it's a secondhand report, not a primary leak, and we only have one side of the employee-vs-company dispute.