strata
Reads Drizzle schema to identify entities, fields, and ID columns for composing schema-aware backend modules.
Generates and verifies backend modules targeting Express, with correct middleware ordering enforced by rank.
Reads Mongoose schemas to identify entities, fields, and ID columns for composing schema-aware backend modules.
Reads Prisma schema to identify entities, fields, and ID columns for composing schema-aware backend modules.
Reads Sequelize schema to identify entities, fields, and ID columns for composing schema-aware backend modules.
Reads TypeORM schema to identify entities, fields, and ID columns for composing schema-aware backend modules.
Your agent writes the backend. Strata proves it runs.
An MCP server that composes verified backend modules into your codebase — reading your schema, following your conventions, wiring them in the order Express actually requires — then writes one command that boots the app and exercises every requirement against a live server.
$ # your agent calls one tool, once
strata_use dir=./shop-api task="product list API"
capabilities=[ "cursor pagination with sorting",
"per-IP rate limiting",
"structured request logging" ]
FILES CREATED
server.js
strata/lib.js — the implementation these import from
strata/verify.js — boots the app and exercises the feature end to end
$ npm install && node strata/verify.js
PASS unit selftests — 3 passed, 0 failed
PASS server boots and answers /health
PASS correlation id honours an inbound x-request-id
PASS an authorization header is NOT written to the log
PASS a password in a request BODY is NOT written to the log
PASS a malformed body is a 4xx and leaks no stack trace to the caller
PASS /items walks pages by cursor without repeating a row
PASS a sort field that is not allowlisted is REJECTED, not honoured
PASS a burst past capacity yields 429 + Retry-After
12/12 checks passed — the delivered feature works end to end.Key capabilities
Schema-aware composition — reads Prisma, Mongoose, Drizzle, TypeORM, Sequelize or plain JS and wires modules against your real entity, fields and ID column
Correct middleware ordering — logging above body parsing, rate limits above routes, error handlers last, enforced by rank rather than left to the model
Generated end-to-end verifier —
strata/verify.jsboots the app on a free port and drives every requirement against itSix machine-checked admission gates — no module reaches your project without passing all of them
Honest declines — refuses roughly a third of tasks, where composing costs more than writing the code
$ # asked for something the library does not cover
strata_use task="slugify helper" capabilities=["convert a string to a url slug"]
No verified Strata recall covers "slugify helper". Build it from scratch the
normal way — a clean hand-written implementation is the right outcome here,
not a forced match.Local by construction — your source and schema never leave the machine; only the task text is sent
The numbers
without Strata | with Strata | ||
tokens | 950,011 | 319,604 | −66% |
turns | 31.3 | 15.0 | −52% |
wall-clock | 190s | 63s | 3× faster |
cost | $0.190 | $0.077 | −59% |
checks passed | 70.8% | 100% |
One backend task — a product API with pagination, per-IP rate limiting and request logging. Claude Haiku 4.5, three runs per arm, mean. Every number is lower and the quality is higher.
Tokens are the whole session: input, output and the cached context re-read on every turn. Across all 18 runs in this battery, 98–99% of a session's tokens are that re-read context — output is under 2%. So output length is not the lever; turns are, and fewer turns is the same thing as fewer tokens is the same thing as less money.
Related MCP server: MCP SSDLC Security Toolkit
What the failed checks actually were
A score is easy to wave away. These are the failures themselves, re-graded from the archived trees. Every one is code that runs, answers 200-or-201, and looks finished.
A malformed request returns your stack trace
One request with a truncated body. Both apps answered 400 — only one of them is safe.
<!DOCTYPE html>
<html lang="en"><head><title>Error</title></head>
<body>
<pre>SyntaxError: Unexpected end of JSON input
at JSON.parse (<anonymous>)
at parse (C:\Users\...\node_modules\body-parser
\lib\types\json.js:96:19)
at C:\Users\...\body-parser\lib\read.js:128:18
at AsyncResource.runInAsyncScope (node:async_...HTML from a JSON API, the parser's internals, and absolute paths from your server's filesystem — handed to whoever sent the bad byte.
{
"error": "malformed JSON in request body",
"details": [
{ "field": "body",
"message": "could not be parsed as JSON" }
]
}The same 400, in the same envelope as every other error, telling the caller what to fix and nothing else.
Failed in 6 of 6 unaided runs, across both tasks. Passed in 6 of 6 with Strata. The grader records it as LEAKS STACK TRACE; nothing in the session's own output mentions it.
A retried order with a different body was accepted anyway
POST /orders Idempotency-Key: k-1 {"items":[ A ]} → 201 Created
POST /orders Idempotency-Key: k-1 {"items":[ B ]} → 200 OK ← order A returnedThat is the subtle half of idempotency, and the half a naive implementation misses entirely. The client asked for a different order and was told its request succeeded. Nothing errors, nothing logs. Order B simply never exists, and the caller holds a 200 saying it does. The correct answer is 409 or 422.
Failed in 3 of 3 unaided runs. Passed in 3 of 3 with Strata.
The API ignored the page size it was asked for
GET /products?limit=5 → 200 OK, 10 itemsPagination that returns whatever it likes. Nothing errors, nothing logs, and the bug reaches whoever consumes that endpoint. One unaided run in three.
The database schema was edited, unasked
The task was "if a client retries the same order request it should not create two orders." It never mentions the data model. One unaided run in three rewrote prisma/schema.prisma; no Strata run touched it.
The wider problem is that you cannot predict which files come back changed. Across three runs of the same prompt, the unaided arm touched six different files — and only three of them in every run. Strata touched the same ten files in all three runs: an identical footprint, run to run.
Run it three times. Get the same answer three times.
task | without Strata | with Strata |
product API | 63%, 75%, 75% | 100%, 100%, 100% |
idempotent orders | 14%, 71%, 71% | 100%, 100%, 100% |
payments + queue | 0%, 0%, 50% | 0%, 100%, 100% |
Zero variance on the tasks the library covers. Three runs of the same prompt return the same score, three times out of three — against a spread of 26.9 points without it.
That 14% is not a grading artefact; it repeats on re-grade. That session invented an order API whose create endpoint rejected every request shape it was sent, and because everything else depends on creating an order, five checks collapsed at once. A cliff, not a slightly worse result — and nothing in the session's own output says it happened.
Where it does not help
Payments is on the board with its failures intact. Both arms shipped a build that does not run: one unaided run never wrote an entry point, and one Strata run pinned bullmq@5.81.3 beside an incompatible redis@4.7.1, which cannot install. Both packages are the model's choice — Strata covers the webhook and nothing else on that task, and the cost ratio lands at 0.95×, a wash.
That is the rule the whole board obeys: the advantage tracks how much of the task the library covers. Where coverage is high the numbers above hold. Where it is one capability out of four, Strata is roughly free and roughly neutral.
Strata also declines outright when a task is below the point where composing beats writing — about a third of the time.
Method
Checks were written from the task prompt alone and frozen before the first run. Every check has a negative control proving it can fail. Grading is a separate suite — never strata/verify.js, which Strata generates and which would be marking its own homework. Every output tree is archived.
n=3, Claude Haiku 4.5, one model per cell. Nothing here speaks to Sonnet or Opus. Cost and token figures move with the model and the prompt; the consistency figures do not.
Full method, per-run scores and every instrument defect found along the way: docs/BENCHMARK.md.
Quick start
Prerequisites: Node.js ≥ 18 and any MCP client — Claude Code, Cursor, Windsurf, VS Code or Claude Desktop.
// .mcp.json (or claude_desktop_config.json for Claude Desktop)
{
"mcpServers": {
"strata": { "command": "npx", "args": ["-y", "stratalib"] }
}
}Restart the client and ask for a backend feature that needs several parts:
Add cursor pagination, per-IP rate limiting and request logging to the products API.Strata reads the project, composes the modules, writes the files, and prints what it created and what it modified. Then:
npm install && node strata/verify.jsNo API key and no account. Modules are served from the hub; the task text is the only thing sent. Your source, schema and files stay on your machine.
The tool
Strata registers exactly one tool. Every tool in an MCP schema is billed on every turn, so the surface is kept to one that does the whole job.
strata_use
Argument | Purpose |
| Absolute path to the project root — where the schema and conventions are read from |
| A short label for the work |
| 3–6 phrases naming the parts of the job. Your model writes these; it has read the whole task |
Returns the files created and modified, the exports available from each module, and the command to verify the result.
How it works
1 · Reads the project — locates the ORM and extracts the real entity: fields, types, enums and the actual ID column. Deterministic, in Node, before the model sees a byte. Where the entity cannot be identified with confidence, Strata leaves a slot rather than guessing.
2 · Selects modules — each capability phrase is scored against the library, and anything matching on shared vocabulary alone is discarded. Fewer than two surviving modules triggers a decline.
3 · Composes — modules contribute to the app rather than owning it, each contribution carrying a rank that fixes its position in the middleware chain. A malformed request throws during body parsing, so logging mounts above it; get that backwards and the one request most worth tracing is the one that loses its correlation id.
4 · Writes the verifier — strata/verify.js runs each module's own suite, boots the app on a free port, and exercises every requirement against it. Built against your entity, so the checks run on your fields and your routes.
Admission gates
Every module passes six machine-checked gates before it can be served. A module that fails is discarded, not repaired — hand-patching generated modules returns coverage to craft and stops it scaling.
Gate | Requirement |
Exports | Loads, and every export it declares resolves at runtime |
Selftest | Its own suite passes, with a stable assertion count across five runs |
Adversarial | ≥ 8 assertions, hostile inputs, and assertions that something must not happen |
Compose | Valid fragments with ranks, and declared factories that exist |
Collisions | No exported name collides with another module |
Composed boot | Composes with two others into an app that starts and verifies |
The adversarial gate is the one that matters. Every hand-written module in this library shipped with a real bug its own tests did not catch — a 404 that reset a circuit breaker's failure count, a dropped enum constraint, an attacker-controlled request id echoed into a response header. A confirmatory suite admits exactly those.
Repository layout
Path | Contents |
| MCP server: project reading, selection, composition, verifier generation |
| CLI entry point |
| Express skeleton used during composition |
| Pre-registered check suites, negative controls, run records and archived output trees |
| Admission gates, library indexing, selection tests |
Modules are served from the hub; the task text is the only thing sent. Your source, schema and files stay on your machine.
Documentation
Document | Subject |
The full benchmark: method, per-run scores, and every instrument defect found | |
What shipped in each release |
Development
npm install
node --max-old-space-size=8192 node_modules/typescript/bin/tsc -p tsconfig.mcp.json # build
node scripts/admit-recall.js recalls/<domain>/<name>/v1 # run the gates
node benchmark/quality/negative-control.js # prove the checks can fail
node benchmark/run-quality-battery.js --tasks catalog --max 3 # collect runsSTRATA_MODE=local composes against a local recalls/ checkout instead of the hub — required when testing a module that has not been deployed.
Acknowledgements
Built on the Model Context Protocol, Express, Prisma, Mongoose, Drizzle, TypeORM and Sequelize.
License
AGPL-3.0-or-later. See LICENSE.
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Connectors
Compile full-stack apps from one architecture spec, with a shared Memory Fabric.
Deploy full-stack web apps with database, file storage, auth, and RBAC via a single API call.
Deterministic context layer for your codebase: change impact, blast radius, answers with receipts.
Build a full backend from Claude Code — boards, data, REST APIs — plus a ready-made admin UI
Related MCP Servers
- FlicenseNot gradedqualityDmaintenanceAn MCP-powered compliance copilot for SaaS stacks, enabling structured audit workflows including stack detection, module wiremapping, implementation directives, code verification, and security/infrastructure/legal readiness gates.
- AlicenseBqualityDmaintenanceEnables orchestrating secure software development pipelines with domain-specific compliance (HIPAA, PCI-DSS, etc.), generating pseudocode, threat models, and CI/CD from user stories via natural language.17MIT
- FlicenseNot gradedqualityDmaintenanceMulti-stack scaffolding engine for generating production-ready module structures (NestJS, Java Spring, Python FastAPI) directly from an MCP-compatible AI client, reducing token usage for boilerplate code.
- AlicenseAqualityAmaintenanceTopos scores code quality by analyzing the geometric and topological structure of program graphs, surfacing structural debt that conventional linters can't compute. Gives coding agents a medal-scored (SLOP → GOLD) feedback loop for writing cleaner, more composable code.1824BSD 3-Clause
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/stratalib/strata'
If you have feedback or need assistance with the MCP directory API, please join our Discord server