Agent output hallucination smell

Judges: claim sentence plus cited tool log snippet. Questions: choice:support - supported/unsupported/contradicted. Action: Strike unsupported claims.

AA clean typed decision (yes/no, a pick from a list, or a level on a scale) that is concrete, repeats routinely, and has meaningful value.3.90

Key facts

Vertical
Software & tech
Function
Agents & dev
Status
Idea
Volume
routine
Value
meaningful
Risk
moderate
Evidence
—
Flags
None

Build this with a classifier

Define a typed decision with a bounded answer, then evaluate it on examples.

{
  "decision_type": "choice",
  "question": "Does this input match the decision in “Agent output hallucination smell”?",
  "input": "<input to classify>",
  "output": "one label from a fixed list"
}

Related use cases

Cite this

Copy a link in your preferred format.