WF-72HNWX
Facebook's Automated Tools Failed to Adequately Remove Hate Speech, Violence, and Incitement
Facebook's automated moderation tools were shown by internal documents performing incomparably to human moderators, and accounting for only a small fraction of hate speech, violence, and incitement content removal.
- Date it happened
- 2021-03-01
- Organisation involved
- Product, system or model
- Facebook's automated moderation tools, powered by artificial intelligence, are intended to flag and remove posts containing hate speech and other detrimental content.
Where this came from
Attribution
This incident was imported from AI Incident Database and is used under CC BY-SA 4.0. Our additions to it — the structured fields, the translation, the checks against other reports — are published under the same licence.
This is a record of what was reported, not a finding that anyone broke the law. If it names your organisation and you believe it is wrong, the corrections process is free and open to everyone.