What content moderation actually involves
Launch library · evergreen read

Content moderation refers to the processes platforms use to review and manage content that may violate their policies, combining automated detection systems with human reviewers who make judgement calls on more complex or genuinely borderline cases. Neither approach works particularly well entirely on its own, which is exactly why most major platforms rely on some combination of the two working together.
Automated systems can quickly identify clearly prohibited material at considerable scale, but they generally struggle with context, sarcasm or nuance, which is why human review remains an important part of the process for anything ambiguous or genuinely contested. This combination is imperfect, and mistakes do occur on both sides of that equation from time to time.
Moderation systems are continually refined as new patterns of harmful content emerge across different platforms, though platforms differ considerably in how transparent they are about their processes overall. This is one reason users are encouraged to use in platform reporting tools whenever they encounter concerning content directly, rather than assuming automated systems will catch everything on their own.