list_hub_models
List Hugging Face instruct/chat models for local inference, with optional query and per-model VRAM estimates.
Instructions
List Hugging Face Instruct/chat safetensors for Late infer on your computer (same searchable catalog as the GUI). Always includes well-known Qwen/Gemma/Mistral/Llama/Phi Instruct ids; Hub API rows merge on top when reachable. Each row includes estimated VRAM at max usage (weights plus KV cache). Optional query (Gemma, Qwen, …). No filesystem path. Does not download. Public Hub ids are listed without Cloud AI; HF_TOKEN from Settings is used when present so gated listings appear. Extra pull_late_infer / delete_late_infer still wait for Approve.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| q | No | Optional search string, e.g. Gemma. Empty lists all families (newest first). |