vram-mcp
Related Servers
Alternatives to vram-mcp
No user-submitted related servers found.
Related Servers
- AlicenseAqualityBmaintenanceEnables AI agents to search, compare, and right-size Hugging Face models, including estimating whether a model fits in available GPU VRAM before downloading it.521 PyPIMIT
- AlicenseAqualityCmaintenanceExposes NVIDIA GPU metrics (info, utilization, VRAM, temperature) via MCP tools for real-time querying from AI assistants.543 PyPIMIT
- AlicenseAqualityCmaintenanceEnables AI agents to search Vast.ai GPU marketplace, deploy compute-heavy video models or ComfyUI, and guarantee instances are torn down after the job finishes.171MIT
- AlicenseBqualityDmaintenanceEnables complete local Ollama management including listing models, chatting with local LLMs, starting/stopping the server, and getting intelligent model recommendations for specific tasks through natural language commands.94MIT
- AlicenseAqualityDmaintenanceEnables AI assistants to manage LM Studio models, including listing, loading, and unloading models through the LM Studio API.628 npm4ISC
- FlicenseNot gradedqualityBmaintenanceEnables AI coding agents to run policy-constrained Python jobs on lab NVIDIA GPU hosts, including GPU inspection, device reservation, job launching and monitoring, and owner-scoped process stopping.2-
TDQS
Scored across 13 tools
Each tool targets a distinct operation: claim/renew/release manage claims, reserve handles capacity reservations, unload/ensure_free/warm manage model residency, while vram_status/list_loaded/list_claims/history/trend provide distinct read views. Even similar actions like unload vs ensure_free are clearly separated by scope (single model vs threshold-based eviction).
Naming uses a mix of bare verbs (renew, release, advise, unload, warm, claim, reserve), verb_noun snake_case (list_claims, list_loaded, ensure_free), and noun-only identifiers (vram_status, history, trend). The pattern is somewhat predictable by category (actions vs queries), but it lacks a single consistent convention, making it less uniform than an all-verb_noun set.
13 tools is well within the ideal 3–15 range and each tool addresses a meaningful aspect of VRAM management: claims, reservations, eviction, warming, status, history, and trend. No tool feels redundant or superfluous.
The surface covers the core lifecycle well: claim/renew/release, reserve, warm/unload/ensure_free, plus status, history, and trend. A minor gap is the lack of a dedicated list_reservations tool—list_claims seems focused on model claims, so capacity reservations may be invisible unless they appear in status/history.