Provides persistent knowledge graph memory for AI agents with local semantic search using Neo4j and ONNX embeddings, enabling offline operation with zero API costs.
High-performance MCP response caching server that reduces latency from ~3000ms to ~0.001ms for repeated tool calls using a dual-layer SQLite and LRU memory cache.