pipecat-docs-mcp
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@pipecat-docs-mcphow to use daily transport in a pipeline"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
pipecat-docs-mcp
Search 317 pages of Pipecat docs + 1,028 GitHub Issues from Claude Desktop or Cursor — grounded answers, no hallucinated APIs.
What it does
Indexes 8,817 chunks from the Pipecat docs and GitHub Issues, then exposes them through 4 MCP tools. Claude can answer questions about configuration, code examples, architecture concepts, and provider comparisons — with source-cited results pulled directly from the official docs and community issue threads.
Retrieval pipeline: BM25 + FAISS dense vectors → Reciprocal Rank Fusion → Cross-encoder reranker
Metric | Value |
Pages indexed | 317 |
GitHub Issues indexed | 1,028 |
Total chunks | 8,817 |
Mean Reciprocal Rank | 0.863 |
Recall@5 | 100% |
Avg query latency | 35ms |
Related MCP server: github-rag-mcp
Tools
Tool | Use it when... |
| General questions, config options, error messages |
| You need a runnable Python pipeline |
| "What is a Frame / Pipeline / VAD / Transport?" |
| Choosing between STT, TTS, or transport providers |
Quickstart
1. Clone and install
git clone https://github.com/DTufail/pipecat-docs-mcp.git
cd pipecat-docs-mcp
python -m venv .venv && source .venv/bin/activate
pip install -r requirements.txt2. Build the search index
data/chunks.jsonl is included. This builds the BM25 and FAISS indexes (~2 min):
python indexer.py3. (Optional) Index GitHub Issues
Adds 1,028 issues from pipecat-ai/pipecat for error-driven retrieval:
export GITHUB_TOKEN="your_token_here" # optional but recommended (5k req/hr vs 60)
python github_indexer.py
python indexer.py # rebuild indexes with issues merged4. Connect to Claude Desktop
Add to ~/Library/Application Support/Claude/claude_desktop_config.json:
{
"mcpServers": {
"pipecat-docs": {
"command": "/absolute/path/to/.venv/bin/python3",
"args": ["/absolute/path/to/pipecat-docs-mcp/server.py"]
}
}
}Use the Python binary from your virtual environment, not system Python. Restart Claude Desktop after saving.
5. Connect to Cursor (HTTP)
fastmcp run server.py:mcp --transport http --port 8000Point your MCP client at http://localhost:8000/mcp.
Project structure
pipecat-docs-mcp/
├── server.py # FastMCP server — 4 MCP tools
├── retrieval.py # Hybrid search (BM25 + FAISS + RRF + reranker)
├── indexer.py # Builds BM25, FAISS, and metadata indexes
├── github_indexer.py # Fetches GitHub Issues → data/github_issues.jsonl
├── config.py # Paths and model config
├── test_server.py # In-process tests for all 4 tools
├── test_retrieval.py # Retrieval eval — MRR, Recall@5, latency (17 queries)
├── requirements.txt
├── pipecat-scraper/
│ ├── scraper.py # Sitemap-driven scraper (317 pages, resumable)
│ ├── utils.py # HTML parsing and chunking
│ ├── audit.py # Per-page extraction audit
│ └── config.py
└── data/
├── chunks.jsonl # 6,936 scraped doc chunks (included)
├── chunk_metadata.json # Title, section, URL, content type per chunk
├── bm25_index.pkl # Generated by indexer.py (gitignored)
├── faiss_index.bin # Generated by indexer.py (gitignored)
├── embeddings.npy # Generated by indexer.py (gitignored)
└── github_issues.jsonl # Generated by github_indexer.py (gitignored)Re-scraping the docs
To refresh against a newer Pipecat release:
cd pipecat-scraper
python scraper.py # scrapes docs.pipecat.ai → data/chunks.jsonl
cd ..
python indexer.py # rebuilds indexesRoadmap
Auto-refresh on new Pipecat releases via GitHub webhook
Expanded coverage — LiveKit and Vapi docs
License
MIT
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- Alicense-qualityCmaintenanceEnables searching and retrieving official PipeCD documentation through full-text search and document retrieval. Clones PipeCD docs from GitHub and indexes Markdown files for quick access to titles and content.6Apache 2.0
- Alicense-qualityAmaintenanceEnables AI agents to search and retrieve context from GitHub issues, pull requests, releases, and documentation using hybrid semantic search and time-ordered activity scans.777Apache 2.0

notifly-mcp-serverofficial
Alicense-qualityCmaintenanceEnables AI agents to search Notifly documentation and SDK code examples for seamless integration.181Inno Setup- AlicenseAqualityCmaintenanceEnables searching and retrieving Reflex documentation, including full-text search, code examples, error analysis, changelog, migration guides, API reference, component props, and recipes.143MIT
Related MCP Connectors
Search @imqueue docs and scaffold typed services & clients from your AI coding agent.
Search public open-source code, documentation, metadata, vulnerabilities, changelogs, and examples.
Search your knowledge bases from any AI assistant using hybrid RAG.
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/DTufail/pipecat-docs-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server