# Content moderation unsafe tagging

Multi-label classification of user content for moderation \(hate speech, spam, graphic violence, sexual content\).

Grade: S — Top tier. Meets A, and the payoff is major with high confidence.. Score: 5.9.

Vertical: Software & tech.

Function: Social & creator.

Status: Seen in the wild.

Volume: high.

Value: major.

Risk: high.

Evidence: described plan.

Flags: human-review.

- [Source](https://arxiv.org/pdf/2407.10995)
- [Vertical](/verticals/saas_tech/)
- [Function](/functions/social_creator/)
- [Social media spam and abuse classification](/use-cases/social-media-spam-and-abuse-classification/)
- [Escalation priority for trust-and-safety reports](/use-cases/escalation-priority-for-trust-and-safety-reports/)
- [How LinkedIn's AI Stops Viral Spam Before Millions See It](/use-cases/how-linkedin-s-ai-stops-viral-spam-before-millions-see-it/)
- [Detecting unwanted content using machine-assisted content moderation](/use-cases/detecting-unwanted-content-using-machine-assisted-content-moderation/)
- [The 60-Second Math Behind Every Spam Filter](/use-cases/the-60-second-math-behind-every-spam-filter/)
- [firehose-judge](/use-cases/firehose-judge/)
