recommend_local_models
Identify local vLLM models compatible with your GPUs using fit flags, and select the newest suitable entry when multiple variants exist. See only CPU-feasible options without an accelerator.
Instructions
Every catalog model for local vLLM, with fit flags for the GPUs on this computer (fits / needs tensor parallel / too big). Newest Hub id is marked when a family has several names (Qwen3.8 over Qwen2.5, Gemma 4 over Gemma 2/3, Llama 4/3.3 over 3.1). Nothing is hidden. Without an accelerator, only tiny CPU-feasible entries fit.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||