vLLM server job
ris_submit_vllm_jobServe a model with vLLM on a full-H100 node and get tunnel instructions. Submit Slurm jobs on the WashU RIS cluster for OpenAI-compatible inference.
Instructions
Serve a model with vLLM (OpenAI-compatible API) on a full-H100 node and return tunnel instructions.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| cpus | No | ||
| port | No | ||
| dtype | No | auto | |
| memGb | No | ||
| model | Yes | ||
| dryRun | No | If true (default) build+validate and return the confirmation summary; DO NOT submit. | |
| account | No | ||
| confirm | No | Must be true (with dryRun=false) to actually submit. | |
| profile | No | Named profile from ~/.risbridge-mcp/config.json. Omit for the default. | |
| project | Yes | Project name; becomes one directory under the storage workspace. | |
| condaEnv | No | ||
| gpuCount | No | ||
| walltime | No | 01:00:00 | |
| maxModelLen | No |