litellm-mcp
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@litellm-mcpUse gpt-4 to explain quantum computing in simple terms"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
litellm-mcp
MCP server for the LiteLLM proxy API. No clone, no build — just npx.
Cursor config
{
"mcpServers": {
"litellm": {
"command": "npx",
"args": ["-y", "@weeebdev/litellm-mcp"],
"env": {
"LITELLM_BASE_URL": "https://litellm-api.up.railway.app",
"LITELLM_API_KEY": "sk-your-key-here"
}
}
}
}From GitHub (no npm):
{
"mcpServers": {
"litellm": {
"command": "npx",
"args": ["-y", "github:weeebdev/litellm-mcp"],
"env": {
"LITELLM_BASE_URL": "https://litellm-api.up.railway.app",
"LITELLM_API_KEY": "sk-your-key-here"
}
}
}
}Local dev (from a cloned repo):
{
"mcpServers": {
"litellm": {
"command": "npx",
"args": ["-y", "."],
"cwd": "/Users/adil/projects/litellm-mcp",
"env": {
"LITELLM_API_KEY": "sk-your-key-here"
}
}
}
}Related MCP server: mcp-consultant
Configuration
Variable | Description | Default |
| LiteLLM proxy URL |
|
| Bearer token / virtual key | (required) |
Tools
LLM: litellm_chat_completion, litellm_completion, litellm_embeddings, litellm_image_generation, litellm_responses_create
Models & health: litellm_list_models, litellm_get_model_info, litellm_health_check, litellm_test_model_connection
Admin: litellm_list_keys, litellm_get_key_info, litellm_generate_key, litellm_list_teams, litellm_get_spend_logs, litellm_list_mcp_servers, litellm_list_mcp_tools
Escape hatch: litellm_api_request — call any endpoint from the Swagger docs
Publish to npm
npm publish --access publicThen anyone can run:
npx -y @weeebdev/litellm-mcpThis server cannot be deployed
Maintenance
Related MCP Connectors
An MCP server that provides an API to LLMs to manage their JumpCloud resources.
MCP server for AI dialogue using various LLM models via AceDataCloud
Focused MCP server for OpenAI image/audio generation (v2.0.0). Wraps endpoints via HAPI CLI.
MCP server providing access to the Scorecard API to evaluate and optimize LLM systems.
Related MCP Servers
- AlicenseBqualityCmaintenanceAn MCP server that provides LLMs access to other LLMs412 npm78MIT
- FlicenseNot gradedqualityDmaintenanceMCP server that interfaces with Gemini and OpenAI CLI tools to enable AI model interactions. It provides a bridge to external AI CLIs with predefined model configurations.-
- FlicenseAqualityCmaintenanceMCP server that connects LLM agents to a local LM Studio instance, enabling model management, OpenAI-compatible chat completions, text completions, and embeddings through a set of tools.91-
- FlicenseNot gradedqualityCmaintenanceA local MCP server exposing the OpenAI platform REST API as tools for file management, fine-tuning, inference, images, audio, batch processing, and organization usage/costs.-