Skip to main content
Glama
Nikitzu

metamind-vault-rag

by Nikitzu

metamind-vault-rag

一个用于 Markdown 目录的检索引擎。它监视文件,增量索引,并回答对它们的混合搜索查询。

向量存储在 sqlite-vec 中,关键词存储在 SQLite FTS5 中,两者通过倒数排名融合(reciprocal rank fusion)进行融合。嵌入通过 fastembed 的 ONNX 模型在进程内运行,因此无需搭建服务器,也无需持有 API 密钥。通过 rerank 额外选项,还可以使用可选的交叉编码器重新评分层。

安装

uv tool install metamind-vault-rag

Related MCP server: mnemonic

入口点

命令

用途

metamind-vault-rag-watcher

监视目录并索引更改

metamind-vault-rag-indexer

一次性完整重新索引

metamind-vault-rag-http

回环 HTTP 搜索 API

metamind-vault-rag-server

stdio MCP 服务器

metamind-vault-rag-doctor

环境和索引诊断

配置

变量

含义

VAULT_PATH

要索引的目录

VAULT_COLLECTION

集合名称,用于限定索引文件的范围

VAULT_HTTP_PORT

回环搜索 API 的端口

VAULT_STATE_DIR

索引、缓存和日志的写入位置。默认为 ~/.vault-rag

索引写入状态目录,以集合命名,绝不会放在语料库内部。指向不同集合或不同状态目录的两个客户端可以在同一台机器上共存,互不知晓对方。

消费者

由任何希望在不运行服务的情况下进行检索的客户端安装。引擎对提问者是谁不持任何立场:其输出中不提及任何客户端,其环境变量均以 VAULT_ 为前缀,并且不会在状态目录之外写入任何内容。

开发

uv run --extra dev pytest

客户端可以通过 uv tool install --from /path/to/this/repo metamind-vault-rag 指向工作副本而不是发布版本。

许可证

MIT

Available Tools

3 tools
search_vaultC

Semantic search over the Obsidian Knowledge vault.

ParametersJSON Schema
NameRequiredDescriptionDefault
kNo
queryYes

Output Schema

ParametersJSON Schema
NameRequiredDescription
resultYes

TDQS

C2.1/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations are absent, so the description should fully disclose behavioral aspects. It merely states 'semantic search' – which implies a read operation but does not explicitly state side effects, resource constraints, or result handling. No details on output stability, rate limits, or side effects are given. Minimal transparency.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness2/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single sentence, which is concise. However, it is under-specified: it lacks necessary detail for an agent to use the tool effectively. The brevity is not a virtue because it omits critical information, making it more of an under-specification than a well-structured concise entry.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The tool has two parameters and an output schema, but the description gives no information about return format, pagination, or semantics. Since annotations are missing, the description must bear the burden of explaining expected behavior. It fails to provide enough context to use the tool safely, especially for a semantic search that could have variable behavior.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters1/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description is the only source of parameter meaning. The description does not mention 'query' or 'k' at all. The schema lists 'k' with a default but no description, and 'query' without context. This is a complete failure to provide any parameter semantics in the description.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose3/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description states it performs 'semantic search' over the vault, which clearly identifies the action and resource. However, it does not differentiate from sibling tools like 'related_notes' or 'expand_search' – the term 'semantic' hints at a method but not why this one is distinct. Purpose is clear but not enriched with scope or contrast.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

No guidance is provided on when to use this tool versus the siblings. The description only states what it does, not when it is the right choice. There is no mention of use cases, exclusions, or alternative recommendations.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 3 tool updatesv0.10.1
    • First observedexpand_search
    • First observedrelated_notes
    • First observedsearch_vault

TDQS

C2.7/5.0

Scored across 3 tools

Disambiguation5/5

Each tool has a clearly distinct purpose: semantic search, finding related notes, and expanding search via wikilinks. There is no overlap or ambiguity.

Naming Consistency4/5

All tools use snake_case and follow a verb-like pattern, but 'related_notes' uses an adjective rather than a verb, creating minor inconsistency with the verb_noun style of the other two.

Tool Count5/5

With only 3 tools, the set is minimal but well-scoped for the server's focus on vault search and navigation. It avoids superfluous tools while covering essential actions.

Completeness5/5

The server covers the core needs of a RAG vault assistant: searching semantically, exploring relationships, and expanding through linked notes. This is a complete set for its intended purpose.

Maintenance

ActivitySlowing
ResponsivenessNo issues

Related MCP Connectors

Related MCP Servers

  • A
    license
    A
    quality
    A
    maintenance
    A generic Markdown vault MCP server with FTS5 full-text search, semantic vector search, frontmatter-aware indexing, incremental reindexing, and non-markdown attachment support that exposes search, read, write, and edit tools.
    44
    228 PyPI
    34
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    MCP server for on-device hybrid search over markdown knowledge bases, combining BM25, vector embeddings, and LLM reranking with link graph and time decay.
    15 npm
    MIT