When you’ll look at a Trace
- A user reported a bad answer.
- Costs jumped and you want to know why.
- Latency spiked and you’re hunting the slow step.
- You’re validating a new version and want to spot-check runs.
- You’re building a dataset from real production behavior.
What’s in a Trace
Inspect a run
The main view — timeline of everything that happened.
Tool calls
What each tool was called with, what it returned.
Grounding
Which parts of the answer are backed by retrieved evidence.
Cost & latency
Per-step cost and time.
Feedback
Add corrections and thumbs from the Trace itself.
Search & filter
Find runs by any criterion.