engine_request_metrics
Reads inference engine metrics including TTFT, TPOT, end-to-end latency, and generation token totals for performance monitoring.
Instructions
[READ] TTFT / TPOT / e2e latency + generation-token totals (where the engine exposes them).
Args: target: Inference target name from config; omit for the default.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| target | No |