run_inference_profile
Execute a diagnostic-only profiling window against a managed vLLM server to capture performance traces for audit and comparison without modifying server state.
Instructions
Run one diagnostic-only profile window against a managed vLLM server.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| plan_token | Yes | ||
| expected_plan_id | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||