Drift evaluation report
drift_reportAggregate a drift run to compute per-ticket fidelity scores, verdict counts, drift rate, and Wilson confidence interval. Lists drifted tickets worst-first and pending unscored items.
Instructions
Aggregate a drift run: per-ticket scores, mean fidelity, verdict counts, drift rate, and — for sampling — a 95% Wilson confidence interval on the true drift fraction extrapolated to the whole Done population. Lists the flagged (partial/drift) tickets worst-first with their gaps, and any pending (unscored) tickets.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| runId | No | Defaults to the latest run. | |
| project | Yes |