engine_request_metrics
Retrieves inference latency metrics including TTFT, TPOT, end-to-end latency, and generation token totals for monitoring engine performance.
Instructions
[READ] TTFT / TPOT / e2e latency + generation-token totals (where the engine exposes them).
Args: target: Inference target name from config; omit for the default.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| target | No |