# How do I pick the confidence threshold for letting a decision run automatically?

There's no single number for a whole system. TypeSafe's own guidance is to gate different actions at different levels depending on what a wrong call costs there, a mistake on an FAQ answer and a mistake on a refund don't deserve the same bar, and to set each threshold by plotting confidence against accuracy on your own labeled data rather than picking a round number like 0.8 in advance.

The pattern is simple once a threshold exists: send the confident cases straight through and route the rest to a person or a more expensive reasoning model, something like `if answer.confidence < 0.8: route_to_human_review(ticket)`. Checking whether that confidence is calibrated in the first place has to happen before any threshold means anything, since a 0.8 that's actually right 60% of the time isn't a safe cutoff wherever you set it.

Start conservative and adjust as you observe real results. [Tessary's own classifiers escalate the same way](/answers/tessary-classifiers/what-happens-when-a-tessary-classifier-flags-a-trace): crossing a threshold writes a finding for triage rather than acting on its own, because the cost of a wrong automatic call is higher than the cost of a short delay.

---

Sources:
- TypeSafe docs, How to build with TypeSafe: https://docs.typesafe.ai/concepts/how-to-build-with-system-one.md (fetched 2026-09-21)
- TypeSafe docs, Confidence: https://docs.typesafe.ai/confidence.md (fetched 2026-09-21)

Source: https://tessary.ai/answers/system-one-models/how-do-i-pick-a-confidence-threshold-for-a-decision
More on System one models: https://tessary.ai/answers/system-one-models
From Tessary, agent reliability for AI agents in production: https://tessary.ai
