Comment toxicity and harassment detection

Detect and flag toxic, abusive, or harassing comments for removal or moderation

STop tier. Meets A, and the payoff is major with high confidence.5.79

Key facts

Vertical
Cross-industry
Function
Social & creator
Status
Seen in the wild
Volume
high
Value
major
Risk
high
Evidence
described plan
Flags
human-review

Source: https://docs.ggwp.com/ggwp-discord-bot

Build this with a classifier

Define a typed decision with a bounded answer, then evaluate it on examples.

{
  "decision_type": "yes_no",
  "question": "Does this input match the decision in “Comment toxicity and harassment detection”?",
  "input": "<input to classify>",
  "output": "yes | no"
}

Related use cases

Cite this

Copy a link in your preferred format.