Skip to main content
Glama

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
HTTP_PORTNoPort for HTTP transport3000
QDRANT_URLNoQdrant server URLhttp://localhost:6333
COHERE_API_KEYNoCohere API key
OPENAI_API_KEYNoOpenAI API key
TRANSPORT_MODENo"stdio" or "http"stdio
VOYAGE_API_KEYNoVoyage AI API key
CODE_BATCH_SIZENoNumber of chunks to embed in one batch100
CODE_CHUNK_SIZENoMaximum chunk size in characters2500
CODE_ENABLE_ASTNoEnable AST-aware chunking (tree-sitter)true
EMBEDDING_MODELNoModel nameProvider-specific
CODE_CHUNK_OVERLAPNoOverlap between chunks in characters300
CODE_CUSTOM_IGNORENoAdditional ignore patterns (comma-separated)
CODE_DEFAULT_LIMITNoDefault search result limit5
EMBEDDING_BASE_URLNoCustom API URLProvider-specific
EMBEDDING_PROVIDERNo"ollama", "openai", "cohere", "voyage"ollama
PROMPTS_CONFIG_FILENoPath to prompts configuration JSONprompts.json
EMBEDDING_RETRY_DELAYNoInitial retry delay (ms)1000
CODE_CUSTOM_EXTENSIONSNoAdditional file extensions (comma-separated)
EMBEDDING_RETRY_ATTEMPTSNoRetry count3
EMBEDDING_MAX_REQUESTS_PER_MINUTENoRate limitProvider-specific

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Server capabilities have not been inspected yet.

Tools

Functions exposed to the LLM to take actions

NameDescription

No tools

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources