darwin-memo
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@darwin-memofind memories about optimizing storage costs"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
darwin-memo
Measured memory for coding agents. Keep repository lessons, connect them to reported task outcomes, and inspect why each lesson stays or is removed.
Start with Python projects using pytest, GitHub Actions, and an MCP client. The core uses only the Python standard library. MCP support is an optional extra.
The product goal is lower total operating cost while preserving task success. That benefit is not established by the committed experiments. A smaller memory store or a larger energy balance is not a savings measurement.
See the mechanism in 60 seconds
From this reviewed checkout:
python -m pip install -e .
darwin-memo demo --out /tmp/darwin-demo.json
darwin-memo ui /tmp/darwin-demo.jsonThe demo uses a filesystem-backed simulation: file deletion uses file size, and protected-file recreation incurs a modeled penalty of three times that size. It does not measure restoration scratch space, I/O cost, or money. Open a lesson in the dashboard to inspect its settlements and removal reason.
To try the coding workflow with a real failing evaluation, use the fixed pytest example. Its authentication-retry defect changes one fixed case from failure to success. The example demonstrates settlement; it does not establish that memory caused the fix.
Related MCP server: NeuralVaultCore
Install the development workflow
The init and task commands in this checkout are unreleased. Install them from
the reviewed checkout on Python 3.10 or later:
python -m pip install -e '.[mcp]'Build the dashboard from this checkout before packaging it:
npm ci --prefix ui
npm run build --prefix uiPersistence requires POSIX advisory locks on a local filesystem. Windows and shared network stores are unsupported; the package refuses lockless persistence.
Connect one repository
Run
darwin-memo init --profile github-pytestat your repository root.Configure your MCP client using
.darwin-memo/mcp.example.json.Add a repository lesson and query memory before a task.
Bind the returned ticket to the repository, task, base commit, and fixed evaluation.
Complete the task and commit the bound store with your changes.
Run the generated GitHub Actions workflow against that task commit.
Download its outcome artifact and carry the accepted store forward.
Open the dashboard to inspect the reported outcome and retained lesson.
The GitHub pytest guide includes setup
requirements and the generated operating instructions.
init refuses overwrites and changes no global agent settings. The recommended
MCP configuration exposes query, add, and inspection while reserving settlement
and retention mutations for the host. It reduces accidental self-awarded credit;
it does not protect against a malicious actor with write access to the store,
evaluation, or workflow.
The workflow compares the same base evaluation at both commits. Added passing tests earn no credit. Green-to-green results earn zero. Missing observations or infrastructure failures abstain and leave the ticket pending. Bound settlements record repository, task, commit, run, evaluation fingerprint, and report hashes. Process one task at a time: independent Git snapshots do not merge.
Understand retention
A query retrieves lessons and opens a ticket. A later outcome moves the consulted lessons' bounded energy balances. Ticks charge upkeep; entries that reach the floor are removed, and similar entries can merge. Pinning exempts a lesson from removal and consolidation. The glossary explains the terms.
Settlement associates outcomes with consulted lessons. It does not establish
causation. Passing tests are observable outcomes, not physically conserved
resources. ci, operator, and agent identify reporting paths. Historical
measured labels are caller claims; absent provenance is unknown. None of these
labels authenticates a report by itself.
CLI, dashboard, and MCP mutations lock the full load–modify–save operation. MCP reloads on each call, cached embeddings survive unrelated writes, and stale Python writers are rejected. Retry the whole operation after contention. See the store format and API reference.
Inspect the evidence
Question | Evidence | What it supports |
What happens under corrupted feedback? | A controlled simulated comparison with a modeled restoration penalty, including failure boundaries | |
Does a tuned counter compete? | Negative findings for universal superiority; tuned forgiveness can retain useful memory under small attack budgets | |
Does memory reduce total coding cost while preserving success? | An unanswered question; historical results lack the required complete cost comparison and noninferiority evidence | |
Can users complete and retain the workflow? | Pending external observation; no fabricated adoption counts |
Reconstruct the paper's central table without an installation or network:
python tools/reconstruct_paper.pyThe paper, Attacking the Curator, focuses on the feedback-corruption threat model, curator comparisons, and a reusable reproduction artifact. Historical replays remain replays. Read the claim audit, reproduction paths, and submission status.
Limitations and participation
Memory can be harmful, useful lessons can starve, and an attacker who controls enough outcome reports can control retention. A fixed evaluation can still be incomplete or gameable. The threat model defines the boundary. If no memory performs best in the cost study, the release will report that result.
Other integrations remain available; the GitHub pytest profile is the supported first-release path. Report setup failures with the diagnostic issue template, or choose a bounded contributor task. Release promotion, participant contact, paid experiments, and submission remain separate authorized actions.
This server cannot be deployed
Maintenance
Related MCP Connectors
Cloud-hosted MCP server for durable AI memory
- memnodeOAuthdev.memnode
Persistent, inspectable memory for AI agents with lineage, correction, and a hosted MCP endpoint.
- KogniteOAuthdev.kognite
Hosted agent memory: store, search, and recall facts across sessions from any MCP client.
An MCP memory server. One memory your agents share — across models, devices and apps.
Related MCP Servers
AlicenseNot gradedqualityAmaintenanceMCP server providing cognitive memory tools (remember, recall, think, etc.) for AI agents, enabling forgetting, consolidation, and contradiction detection.175Apache 2.0- AlicenseNot gradedqualityCmaintenanceAn MCP server that provides persistent long-term memory for AI agents via local SQLite storage with low token overhead, enabling memory storage, retrieval, and management across sessions.1MIT
- AlicenseNot gradedqualityAmaintenanceMCP server for brainmem, providing memory search, write, outcome, explain, and status tools to give LLM agents auditable long-term memory. It prioritizes failures, gates writes on surprise, and supports validity intervals for beliefs.Apache 2.0
- AlicenseAqualityAmaintenanceAn MCP server that gives AI agents durable, temporal memory over local markdown vaults, with tools for searching, asserting facts, querying point-in-time state, and reinforcing useful knowledge.15494 npmMIT