Detect policy-violating posts in internal chat

Judges: internal chat message. Questions: noul:possible_harassment; noul:shares_confidential_info. Action: routes to HR/compliance queue for human review.

STop tier. Meets A, and the payoff is major with high confidence.5.39

Key facts

Vertical
Cross-industry
Function
Ops & HR
Status
Idea
Volume
high
Value
major
Risk
moderate
Evidence
—
Flags
None

Build this with a classifier

Define a typed decision with a bounded answer, then evaluate it on examples.

{
  "decision_type": "choice",
  "question": "Does this input match the decision in “Detect policy-violating posts in internal chat”?",
  "input": "<input to classify>",
  "output": "one label from a fixed list"
}

Related use cases

Cite this

Copy a link in your preferred format.