terrorism ends.

Communities over fear.

Online safety

What content moderation actually involves

Launch library · evergreen read

Photo: Ekushey Wiki gathering 2024, Chattogram (165124) by Wikimedia Bangladesh‎ (CC BY-SA 4.0), via Openverse

Content moderation refers to the processes platforms use to review and manage content that may violate their policies, combining automated detection systems with human reviewers who make judgement calls on more complex or genuinely borderline cases. Neither approach works particularly well entirely on its own, which is exactly why most major platforms rely on some combination of the two working together.

Automated systems can quickly identify clearly prohibited material at considerable scale, but they generally struggle with context, sarcasm or nuance, which is why human review remains an important part of the process for anything ambiguous or genuinely contested. This combination is imperfect, and mistakes do occur on both sides of that equation from time to time.

Moderation systems are continually refined as new patterns of harmful content emerge across different platforms, though platforms differ considerably in how transparent they are about their processes overall. This is one reason users are encouraged to use in platform reporting tools whenever they encounter concerning content directly, rather than assuming automated systems will catch everything on their own.

Back to the library

Share

Sharing opens the network in a new tab. No tracking scripts are loaded on this page.

Printed from terrorism ends.. Sources for this article are listed at the end of the page.