Skip to main content
Glama

RAGBuddy

RAGBuddy

A multi-project RAG (Retrieval-Augmented Generation) platform for coding agents and developers. RAGBuddy provides project-aware access to documentation and knowledge through a web dashboard with AI chat, CLI, and MCP server — all powered by the same underlying core.

What it is

RAGBuddy indexes the docs/ folder (configurable) of one or more registered Git repositories into Qdrant, a vector database, and exposes that index three ways:

  • CLI — ragbuddy ingest/sync/sync-all/search/ask/hook/project/mcp/web

  • Web dashboard — register projects, browse indexed files, upload extra documents, search, chat with a project's indexed docs, run ingest/sync with a live log, review sync history, toggle auto-sync, and copy per-project MCP config, all from a browser

  • MCP server — a coding agent working in your repo can call get_project_context for a quick orientation, then search_project_docs to find the architecture doc, feature spec, or issue writeup relevant to what it's doing right now, instead of relying on whatever happened to fit in its context window

Related MCP server: reflens

Features

  • 🧠 Multi-project RAG with project-isolated retrieval

  • šŸ’¬ Web AI chat over project knowledge

  • šŸ”Œ MCP server for coding agents

  • šŸ“š Repository documentation indexing

  • šŸ“Ž Upload PDF, Word, Excel, Markdown, CSV, and text documents

  • šŸ”„ Incremental synchronization using content hashes

  • šŸŖ Git auto-sync on commit, pull, and checkout

  • šŸ—„ļø Single Qdrant collection with project-level isolation

  • 🧩 Ollama and OpenAI-compatible embedding providers

  • šŸ–„ļø Web Dashboard and CLI using the same core implementation

Architecture

RAGBuddy system architecture

Coding Agents / Web Chat
          │
          ā–¼
       RAGBuddy
          │
    ā”Œā”€ā”€ā”€ā”€ā”€ā”“ā”€ā”€ā”€ā”€ā”€ā”
    │           │
   MCP       Web / CLI
    │           │
    ā””ā”€ā”€ā”€ā”€ā”€ā”¬ā”€ā”€ā”€ā”€ā”€ā”˜
          ā–¼
     RAG Pipeline
          │
          ā–¼
       Qdrant

Each project is isolated using a project field in the Qdrant payload. Retrieval operations are always filtered by project.

See docs/steering/architecture.md for the detailed architecture.

Quick Start

Requirements

  • Node.js 18+

  • npm

  • Docker

  • Qdrant

  • Optional: Ollama for local embeddings

Install

git clone <this-repository>
cd ragbuddy

npm install
npm run build

cp .env.example .env

Start Qdrant:

docker compose up -d

Configure Embeddings

For local Ollama:

EMBEDDING_PROVIDER=ollama
EMBEDDING_BASE_URL=http://localhost:11434
EMBEDDING_MODEL=bge-m3

Then:

ollama pull bge-m3

OpenAI-compatible embedding providers are also supported.

See docs/steering/setup.md for configuration details.

Register a Project

ragbuddy project register <id> <repository>

Example:

ragbuddy project register my-project /path/to/my-project

Projects can also be managed from the Web Dashboard.

Index & Sync

Initial indexing:

ragbuddy ingest <project-id>

Incremental synchronization:

ragbuddy sync <project-id>

Install Git auto-sync:

ragbuddy hook install <project-id>

Scheduled re-sync fallback — syncs every registered project, isolating per-project failures; run this periodically from cron/Task Scheduler as a safety net for projects where the git hook was skipped or never installed:

ragbuddy sync-all

RAGBuddy uses the Git repository as the source of truth. Qdrant acts as a rebuildable search index.

Search & Ask

Query a project's indexed docs directly, without opening the dashboard or a chat session:

ragbuddy search <project-id> "<query>"

Or get a one-shot, RAG-grounded answer straight in the terminal (rewrite → hybrid vector+BM25 search → rerank → one LLM completion, same pipeline the web chat uses):

ragbuddy ask <project-id> "<query>"
ragbuddy ask my-project "how does auto-sync work?"

Web Dashboard

RAGBuddy includes a web dashboard for managing projects, documents, RAG search, AI chat, ingestion, synchronization, and MCP configuration.

Dashboard

Dashboard

Project Overview

Project overview

Project Documents

Project documents

Project search

Sync History

Sync history

MCP Setup

MCP setup

AI Chat

AI Chat overview AI Chat with related documents

Start the dashboard:

npm run web

Open:

http://localhost:4300              → landing page
http://localhost:4300/dashboard    → the dashboard itself
http://localhost:4300/login        → login screen (only when dashboard login is enabled)

For frontend development:

cd web
npm install
npm run dev

MCP

RAGBuddy exposes project knowledge through MCP.

Available tools:

Tool

Purpose

get_project_context

Get a compact overview of the current project

search_project_docs

Semantic search over project knowledge

get_project_document

Read a specific document

list_project_knowledge

List indexed project knowledge

Example:

claude mcp add ragbuddy -- node /absolute/path/to/ragbuddy/dist/cli/index.js mcp

The current project can be resolved automatically from the agent's working directory.

See docs/steering/mcp.md for MCP configuration and usage.

Teaching your agent to actually use it

The four tools above are visible to your agent automatically once the MCP server connects — no extra config needed for that. But an agent only reaches for a tool it happens to think of; it won't necessarily call get_project_context before diving into a task just because the tool exists. Add this to the registered project's AGENTS.md / CLAUDE.md so your agent knows when to use each one:

## Knowledge Retrieval Strategy (ragbuddy MCP)

Before implementing a non-trivial feature in this project:
  1. Use `get_project_context` first to understand the project (identity, Git status, tech stack/architecture summaries, doc inventory).
  2. Use `search_project_docs` for architecture, business rules, historical issues, conventions, and documented behavior.
  3. Use `get_project_document` to read a full doc found via search when a snippet isn't enough.
  4. Use `list_project_knowledge` to see everything currently indexed when orienting from scratch.
  5. Read the actual source code before making implementation decisions — treat it as the final authority for current behavior.

Don't force `get_project_context` for trivial tasks where it adds no value.

Integrating with Your Own Web App

Besides the bundled dashboard, IDE/MCP, and CLI, RAGBuddy's REST API can be called directly from a separate web app or chatbot — no need to run your own vector database, embedding model, or retrieval pipeline.

  • Just need RAG context (you already have your own chat/LLM): POST /api/projects/:id/search — returns the most relevant chunks, ranked with the same hybrid vector+BM25 search, query rewriting, and reranking pipeline the dashboard's own chat uses.

  • Don't have a chat feature yet: POST /api/projects/:id/chat — a full RAG-grounded chat endpoint, streamed via Server-Sent Events, with source citations.

curl -X POST http://localhost:4300/api/projects/<project-id>/search \
  -H "Content-Type: application/json" \
  -d '{"query":"How does incremental sync work?"}'

Both endpoints are open by default (local-trust model, same as the rest of the API) — optionally lock them down with an API key (generated from the dashboard's Settings page) and a CORS allowlist before exposing them beyond localhost.

See docs/features/12-external-web-app-integration.md for full request/response formats, JavaScript/curl examples, and hardening options.

Knowledge Sources

RAGBuddy can index:

Repository
ā”œā”€ā”€ README.md
└── configured documentation paths
    ā”œā”€ā”€ features/
    ā”œā”€ā”€ steering/
    ā”œā”€ā”€ issue/
    └── ...

Additional documents can be uploaded through the Web Dashboard.

All indexed knowledge is stored in a single Qdrant collection and isolated using the project identifier.

Documentation

Documentation

Description

docs/steering/

Architecture, stack, setup, routing, system flow, and conventions

docs/features/

Feature documentation

docs/issue/

Issues and root-cause analysis

docs/design-system/

Web UI design system

landing/

Static landing page — served at / by ragbuddy web

Development

npm run typecheck
npm test
npm run build
npm run web

Frontend:

cd web
npm run dev
npm run build
npx oxlint src

See CLAUDE.md and AGENTS.md for the coding-agent workflow.

License

See LICENSE.

A
license - permissive license
Not graded
quality - not tested
B
maintenance

Maintenance

–Maintainers
–Response time
–Release cycle
–Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    A
    quality
    B
    maintenance
    An MCP server that indexes reference repositories and provides tools for AI coding agents to retrieve lossless code context, enabling reasoning over codebases larger than the agent's context window.
    8
    2
    Apache 2.0
  • F
    license
    Not graded
    quality
    B
    maintenance
    Local MCP server that indexes documentation from URLs/files into a vector database, enabling coding agents to search and use up-to-date library and API documentation.

View all related MCP servers

Related MCP Connectors

  • An MCP server that gives your AI access to the source code and docs of all public github repos

  • Agent-native MCP server over the public saagarpatel.dev corpus. Read-only, stateless.

  • Augments MCP Server - A comprehensive framework documentation provider for Claude Code

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/azmirizkifar20/RAGBuddy'

If you have feedback or need assistance with the MCP directory API, please join our Discord server