
Meta employees warn AI moderation rollout is too fast
Quick Answer
Meta plans to replace half of human moderation requests with AI by 2025, aiming for over 90% for specific content types by year-end.
Quick Take
Employees express concerns about the rapid rollout of AI moderation, highlighting potential risks to content quality and user safety.
Key Points
- Meta aims to replace 50% of moderation requests with AI by 2025.
- The goal is to achieve over 90% AI moderation for certain content by year-end.
- Employees warn that the rapid rollout could compromise content quality.
- Concerns include potential risks to user safety and moderation effectiveness.
- are central to Meta's moderation strategy.
📖 Reader Mode
~1 min readMeta has already replaced roughly half of all human moderation requests with large language models in 2025 and plans to push that share above 90 percent for some content types by the end of the year. The shift is expected to save the company billions annually, according to the Financial Times. Meta disputes the cost argument and points to quality instead, saying that since March, tests show its language models make 13 percent fewer errors than humans when enforcing content policies while catching 10 percent more actual violations. Unlike traditional ML classifiers that struggle with satire or evolving language, the language models are supposed to better grasp nuance and cover more languages.
Employees paint a different picture. One insider says the models still remove or shadow-ban harmless content, and there isn't enough oversight for such a rapid rollout. The transition is already leading to layoffs, especially among external contractors.
There's also a model swap happening behind the scenes, the Financial Times reports. Meta had been using Google's Gemini for moderation and support but recently told staff to switch to its own new foundation model called Muse Spark. The models are trained on past decisions made by human reviewers.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

