inference_metrics
Retrieve session-local inference metrics including call count, cloud split, token totals, per-model breakdown, and average latency. Reports delegation usage from prism_infer only.
Instructions
Returns the current session's local-model inference metrics — call count, local vs cloud split, token totals, per-model breakdown, and average latency. Read-only, no arguments. Reflects prism_infer delegation usage only, not the host model's (Claude's) own token spend (use /cost for that).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||