decision_stats
Generate a calibration report that compares confidence buckets against observed accuracy from recorded outcomes, showing whether high-confidence answers are actually correct.
Instructions
Calibration report: confidence buckets vs observed accuracy from recorded outcomes. Shows whether high-confidence answers are actually right that often.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||