run_inference_profile
Run a diagnostic profiling session on a managed vLLM server to collect torch_profiler or Nsight Systems traces for performance auditing.
Instructions
Run one diagnostic-only profile window against a managed vLLM server.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| profiler | Yes | ||
| server_name | Yes | ||
| scenario_name | Yes | ||
| timeout_seconds | No | ||
| expected_plan_id | No | ||
| measurement_run_id | Yes | Successful compatible unprofiled inference run to link. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||