MemoryBuddy
Enables Hermes to store and recall user facts, preferences, and long-term context across sessions via a shared MCP memory server.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@MemoryBuddyremember that I like coffee and prefer dark mode"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
MemoryBuddy ๐ง
Give your AI agents a shared memory that lasts. Deploy once, connect any MCP-compatible AI tool โ Hermes, Trae, Cursor, Claude Desktop โ they all share the same memory.
๐ What is this?
Most AI tools suffer from "goldfish memory" โ refresh the page, start a new session, switch to another app, and everything's gone. You keep reintroducing yourself, re-explaining your preferences, re-stating context.
MemoryBuddy fixes this with a shared memory layer that any AI tool can read from and write to:
๐ง Long-term memory โ facts, preferences, decisions persist across sessions
๐ Semantic search โ find relevant memories by meaning, not just keywords
๐ค Auto fact extraction โ LLM automatically distills what's worth remembering
๐ Smart summarization โ long conversations get compressed, key points retained
๐๏ธ One-click forget โ
DELETEwipes everything, GDPR compliant๐ MCP protocol โ any MCP-compatible client can connect, zero integration code
๐ธ $0/month โ runs entirely on Cloudflare's free tier
Related MCP server: GroundMemory
๐ก What problem does it solve?
๐ฃ Without MemoryBuddy | โ With MemoryBuddy |
Every AI tool starts fresh โ you re-explain yourself constantly | All your AI tools share one memory โ tell one, they all know |
Switching from Hermes to Trae means losing all context | Switch freely โ memory lives in the cloud, not in the tool |
AI forgets your preferences between sessions | Preferences persist forever, across all sessions and all tools |
Long conversations hit context limits | Auto-summarization keeps things compact |
Privacy concerns โ can't delete what it remembers | One API call wipes everything, fully GDPR compliant |
๐๏ธ Architecture
โโโโโโโโโโโ โโโโโโโโโโโ โโโโโโโโโโโ โโโโโโโโโโโ
โ Hermes โ โ Trae โ โ Cursor โ โ Claude โ
โโโโโโฌโโโโโ โโโโโโฌโโโโโ โโโโโโฌโโโโโ โโโโโโฌโโโโโ
โ MCP โ MCP โ MCP โ MCP
โผ โผ โผ โผ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ MemoryBuddy Worker (Cloudflare) โ
โ โ
โ /mcp โ MCP Server (5 tools, Streamable HTTP) โ
โ /chat โ HTTP API (SSE streaming + auto-extract)โ
โ /memory/:userId โ REST API โ
โโโโโโโโโโโโฌโโโโโโโโโโโโโโโโโโโฌโโโโโโโโโโโโโโโโโโโโโ
โ โ
โโโโโโโผโโโโโโ โโโโโโโโผโโโโโโโ
โ D1 (facts)โ โ Vectorize โ
โ SQLite DB โ โ (embeddings)โ
โโโโโโโโโโโโโ โโโโโโโโโโโโโโโThree-tier memory:
Short-term (Durable Object) โ current conversation context
Long-term (D1 database) โ structured facts: name, preferences, key entities
Semantic (Vectorize) โ vector embeddings for meaning-based recall
๐ Quick Start (3 steps, ~5 minutes)
Prerequisites
Cloudflare account (free is fine)
Node.js 18+
1. Clone & Install
git clone https://github.com/Trainspotting31/memory-buddy.git
cd memory-buddy
npm install2. Create Cloudflare Resources
npx wrangler login
# Create D1 database
npx wrangler d1 create memory-buddy-db
# Create Vectorize index
npx wrangler vectorize create memory-buddy-index --dimensions 768 --metric cosine
# Initialize database schema
npx wrangler d1 execute memory-buddy-db --remote --file=schema.sqlCopy the generated database_id into wrangler.toml (rename from wrangler.toml.example).
3. Deploy
npx wrangler deployDone! Your memory server is live at https://memory-buddy.<your-subdomain>.workers.dev ๐
๐ Connect Your AI Tools
MemoryBuddy speaks MCP (Model Context Protocol). Any MCP-compatible tool can connect โ they all share the same memory.
Hermes Agent
hermes mcp add memory-buddy --url https://memory-buddy.<your-subdomain>.workers.dev/mcpTrae IDE
Settings โ MCP โ Add Manually
Type: Streamable HTTP
URL:
https://memory-buddy.<your-subdomain>.workers.dev/mcp
Or create .trae/mcp.json in your project:
{
"mcpServers": {
"memory-buddy": {
"type": "streamable-http",
"url": "https://memory-buddy.<your-subdomain>.workers.dev/mcp"
}
}
}Cursor
Add to ~/.cursor/mcp.json:
{
"mcpServers": {
"memory-buddy": {
"url": "https://memory-buddy.<your-subdomain>.workers.dev/mcp"
}
}
}Claude Desktop
Add to claude_desktop_config.json:
{
"mcpServers": {
"memory-buddy": {
"type": "streamable-http",
"url": "https://memory-buddy.<your-subdomain>.workers.dev/mcp"
}
}
}Any MCP Client (raw config)
Endpoint: https://memory-buddy.<your-subdomain>.workers.dev/mcp
Transport: Streamable HTTP
Auth: None (or add your own)๐ ๏ธ MCP Tools
Once connected, the AI gets 5 tools:
Tool | What it does | When AI calls it |
| Load all memory for a user | Start of conversation |
| Semantic search by meaning | "What did I say about X?" |
| Save a new fact | User shares preferences, decisions |
| Delete all memory | User says "forget everything" |
| List all memory spaces | Checking what exists |
Shared memory: All tools default to userId: "hermes-shared". Use different userIds to isolate memory per project/persona.
๐ก HTTP API (no MCP needed)
POST /chat โ Chat with memory
curl -N -X POST https://your-worker.workers.dev/chat \
-H "Content-Type: application/json" \
-d '{"userId":"user123","message":"Hi! I'm John and I love espresso."}'GET /memory/:userId โ Get all memory
curl https://your-worker.workers.dev/memory/user123DELETE /memory/:userId โ Wipe memory
curl -X DELETE https://your-worker.workers.dev/memory/user123GET /health โ Health check
curl https://your-worker.workers.dev/healthโ๏ธ Configuration
Edit wrangler.toml:
[vars]
LLM_MODEL = "@cf/meta/llama-3.2-3b-instruct" # Default: Workers AI (free)
# Optional: use external LLM instead of Workers AI
LLM_API_KEY = "sk-your-key"
LLM_API_BASE = "https://api.openai.com/v1"
LLM_MODEL = "gpt-4o-mini"๐ธ Why Cloudflare Free Tier?
Component | Free Tier | Self-Hosted Equivalent |
Compute (Workers) | 100K req/day | $5โ$50/mo (VPS) |
Database (D1) | 1GB storage | $10โ$100/mo (Postgres) |
Vector DB (Vectorize) | 256K vectors | $70+/mo (Pinecone) |
LLM (Workers AI) | 10K neurons/day | $10+/mo (API) |
Total | $0 | ~$100+/mo |
๐ Project Structure
memory-buddy/
โโโ src/
โ โโโ index.ts # Hono router: /mcp + /chat + /memory + /health
โ โโโ mcp.ts # MCP Server factory (5 tools, stateless)
โ โโโ agent-do.ts # Durable Object: chat session + memory orchestration
โ โโโ llm.ts # LLM abstraction (Workers AI / OpenAI-compatible)
โ โโโ memory/
โ โโโ extract.ts # LLM-powered fact extraction
โ โโโ retrieve.ts # Hybrid retrieval (D1 + Vectorize)
โ โโโ summarize.ts # Conversation summarization
โโโ public/index.html # Built-in demo chat UI
โโโ schema.sql # D1 database schema
โโโ wrangler.toml.example # Cloudflare config template
โโโ package.json๐ฎ Try the Demo
Open your Worker URL in a browser โ you'll see a built-in chat interface.
Tell the agent your name and a preference ("I'm Sarah, I'm allergic to peanuts")
Refresh the page
Ask: "What do you know about me?"
It remembers everything. That's MemoryBuddy.
๐บ๏ธ Roadmap
MCP Server (Streamable HTTP)
Multi-agent shared memory
Semantic search
Auto fact extraction
Memory categories & filtering
User authentication
Batch memory import/export
Multi-language support
Hermes plugin (auto-inject memory at conversation start)
๐ค Contributing
Fork โ 2. Branch โ 3. Commit โ 4. Push โ 5. PR
๐ License
MIT โ see LICENSE
This server cannot be deployed
Maintenance
Related MCP Connectors
An MCP memory server. One memory your agents share โ across models, devices and apps.
Shared, governed long-term memory for AI agents across tools and sessions via MCP and REST.
Shared long-term memory vault for AI agents with 20 MCP tools.
Persistent AI memory shared across Claude, ChatGPT, coding agents, and compatible MCP clients.
Related MCP Servers
- AlicenseNot gradedqualityAmaintenanceAn MCP server that extends AI agents' context window by providing tools to store, retrieve, and search memories, allowing agents to maintain history and context across long interactions.MIT
- AlicenseNot gradedqualityCmaintenanceAn MCP-native, local-first memory server that gives AI agents persistent, structured memory across sessions and tools, enabling them to maintain identity and context without reconfiguration.3MIT
- AlicenseAqualityCmaintenanceMCP server for long-term agent memory, providing persistent memory, searchable knowledge, and evolving identity for AI agents.53Apache 2.0
- FlicenseNot gradedqualityBmaintenanceA persistent memory server for AI agents using MCP protocol, enabling semantic storage and retrieval of dialogues, documents, and agent states.-