Why can't I just reproduce the failure locally?

A local re-run rarely reproduces a production failure because the original conditions are gone by the time you try it. The message you retype is a paraphrase, not the literal input. A live tool call today hits current state, not the state the agent saw when it failed. A retrieval step returns whatever the index holds now, which may already differ from what it held then. Agent failures are usually context-dependent: the same wording with a different tool result behaves differently. So a local attempt tests a related case, not the one that happened, and a fix that passes it can still fail on the original. You don’t get to hold the world still while you debug, which is exactly why the trace has to capture that context at the time, or it’s lost for good.

keep reading

More on this.

Two ways to run Tessary.

Tessary is an open-source agent reliability platform. Cloud and self-hosted run the same workflow on the OpenTelemetry traces your agent already emits.

Tessary Cloud

We host it for you. Send your first trace with nothing to deploy and no model key.

what's includedper organization
traces
10,000 per calendar month
stored trace data
1 GB
retention
30 days
model credit
$10, one-time, for triage and root-cause analysis
credit card
not required

Self-hosted Tessary

Run the open-source code on your own infrastructure with one command. Add your own model key for triage and root-cause analysis.

Self-host Tessary for me by following https://github.com/tessaryai/tessary/blob/main/setup.md

docker compose -f oci://docker.io/tessaryai/tessary:compose up -d -y