Does checking whether an answer is grounded require the model that generated it?

No. A paper published this month trained a detector, called a Grounding Probe, on the hidden states of a model that never generated the answer at all, an “observer” that only reads the context, the question, and someone else’s response. Across four different observer models, it reached 0.88 to 0.89 AUROC on RAGTruth, a standard test set for this kind of error, and adding a supervised detector alongside it pushed that to 0.92. One probe held up across six different generators, including ones it never saw in training.

That removes a real constraint. Every earlier hidden-state method read the generating model’s own activations, so a closed-weight generator, or a generator that changes, broke the detector. An observer model sidesteps that the same way Tessary’s groundedness classifier already does, by scoring the answer against its source with a separate model rather than interrogating the one that wrote it.

Asking the observer model outright, instead of reading its hidden state, cost at least 0.166 AUROC in every model tested. The hidden state carries more than the model says when asked directly.

sources

keep reading

More on this.

Two ways to run Tessary.

Tessary is an open-source agent reliability platform. Cloud and self-hosted run the same workflow on the OpenTelemetry traces your agent already emits.

Tessary Cloud

We host it for you. Send your first trace with nothing to deploy and no model key.

what's includedper organization
traces
10,000 per calendar month
stored trace data
1 GB
retention
30 days
model credit
$10, one-time, for triage and root-cause analysis
credit card
not required

Self-hosted Tessary

Run the open-source code on your own infrastructure with one command. Add your own model key for triage and root-cause analysis.

Self-host Tessary for me by following https://github.com/tessaryai/tessary/blob/main/setup.md

docker compose -f oci://docker.io/tessaryai/tessary:compose up -d -y