MCP Local LLM Server
by iagomussel
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
| MAX_TOKENS | No | Maximum tokens in response. Default: 256. | |
| MODEL_NAME | No | Model name to use. Examples: 'llama3', 'gpt-3.5-turbo', 'claude-3-haiku-20240307', 'gemini-1.5-flash'. | |
| OLLAMA_URL | No | URL of the Ollama server. Default is http://localhost:11434. | |
| TEMPERATURE | No | Temperature for response generation. Default: 0.7. | |
| LLM_PROVIDER | No | Select provider: 'ollama', 'openai', 'anthropic', or 'gemini'. Default is 'ollama'. | |
| GEMINI_API_KEY | No | API key for Google Gemini. | |
| OPENAI_API_KEY | No | API key for OpenAI. | |
| ANTHROPIC_API_KEY | No | API key for Anthropic. | |
| DISABLE_CHAT_SUMMARY_RULE | No | Set to 'true' to disable the chat_end_summary_rule prompt. |
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Server capabilities have not been inspected yet.
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
No tools | |
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
This server cannot be deployed
Maintenance
ActivityInactive
ResponsivenessNo issues