Core concepts
Traces
A recorded agent run: input, output, steps, cost, and timing.
Steps
A single LLM call or sub-operation within a trace.
Scenarios
A named set of test items run against your agent.
Assertions
Checks that pass or fail a completed run.
Replay
Run your agent against a recorded trace with mocked LLM responses.
Snapshots
A stored known-good run that new runs are compared against.