Read the flow trace behind one scored item
caliper_evals_runs_item_executionWhat the flow actually DID on one run item: every step in order (what it said, which tools it called with what arguments, what came back) and the final answer, plus status, duration, and cost. Read this when a low score needs explaining beyond the judge's reasoning — a wrong tool call or an empty tool result is usually the cause, and the fix is different from a prompt fix. Pass the run item's id (NOT itemId) from caliper_evals_runs_get. Items whose outputs were supplied from outside (external evals) have no trace and return not_found. Steps are capped for transport; the run page in Caliper has the full record.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| runId | Yes | Run id. | |
| evalId | Yes | Eval the run belongs to. | |
| runItemId | Yes | The run item's `id` from caliper_evals_runs_get (not its dataset itemId). | |
| workspace | No | Workspace slug. Personal tokens with no default workspace MUST pass this; tokens with a default can override per call. Ignored for workspace API keys. |