slimtoken.list_model_presets
List recommended local-model presets by GPU VRAM tier (4/8/16GB) with usable context. Optionally measure real token reduction on a bloated payload.
Instructions
List recommended local-model presets by GPU VRAM tier (4/8/16GB), each with a usable context. With measure=true, enriches each row with the live measured token reduction on a bloated payload (run by the pipeline itself).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| measure | No | run the pipeline to measure real reduction | |
| vram_gb | No | filter to one tier (4/8/16) |