"Understanding Inference Models" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Simulate cloud infrastructure and estimate cost before deploying or changing infrastructure.
Live GPU compute & inference token price indices for AI agents — 591 reference indices across H100/A100/B200/B300 spot+on-demand, Claude/GPT/Llama token pricing, and more. Every value is methodology-versioned and citable via the /v1/verify handshake.
Run DeepInfra inference, list models and read account rate limits.
Find where to rent a cloud GPU right now. find_gpu searches 22 providers for H100, H200, B200, B300, RTX 4090 and 70 more models, with the price per GPU hour and the stock each provider reported at the last daily check. Filter by model, node size and max price. Free, no key.
GPU and LLM inference benchmarks, hardware evidence, deployment recommendations, and launch configs.
Cloudflare Workers MCP server: ai-model-router
Remote MCP for RunComfy: ComfyUI deployments, hosted models, LoRA training. 31 tools.
Which open models fit your GPU or Mac, measured. Model momentum, GPU rental prices, weekly pick.
A second opinion for AI agents: one prompt across several live Gonka models + roles, one call.
Signed-in OneInfer models, chat, media, Studio movies, music, GPUs, and credits.
Korean cloud for AI agents — create servers, deploy Docker apps, manage domains, call 300+ AI models
Robot embodied-AI cloud for GPU training, inference, benchmarks, and robot data collection.
Run AI models, create deployments, and manage predictions via cloud API