FailTrace
Provides integration with Git repositories to capture source identity for verification, bisect changes and identify the first-parent boundary where a regression was introduced, and support baseline capture before edits.
Provides support for following exact Unity tests through a fix, using fresh NUnit 3 reports to distinguish failing baselines, passing candidates, and skipped or inconclusive tests, as documented for Windows EditMode examples.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@FailTracecapture this flaky checkout test failure and save a baseline"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
FailTrace
Reproduce the failure. Check the fix.
A test fails intermittently. Your coding agent changes the code. One passing retry leaves you guessing.
FailTrace repeats the same check, saves a failing baseline, and checks the proposed fix against it. Use the CLI or MCP tools to get trial results, a smaller reproducer, and evidence your agent can inspect before accepting a patch.
Local execution. No AI API, account, or telemetry required. Keep your existing test runner and assertions.
Quick start
With Node.js 22.12+ and npm, run this in any working directory:
npx --yes failtrace demo
Static walkthrough · Static poster · Demo guide
The demo reduces six input items to ["BUG"], rejects a patch that crashes for another reason, and checks a working patch. It saves the evidence in .failtrace/ and prints a replay command.
Related MCP server: ResiliReplay
For coding agents
Connect the local stdio MCP server through your client's configuration:
npx --yes failtrace@1.5.0 mcp --cwd "/absolute/path/to/your/project"Copy the MCP configuration and check the connection →
Your client launches this command. Running it alone in a terminal waits for MCP requests. The guide includes Windows setup.
Then ask your agent:
Use FailTrace to capture this test failure before editing. Choose the exact test or failure message, save a bounded baseline, and inspect the matching trial. After the change, verify against that baseline and explain any unrelated errors or incomplete evidence.
The seven MCP tools share the CLI's Core engine. They retain the failure signature and investigation evidence across repetition, comparison, regression search, minimization, verification and replay. Agents can retrieve saved trial and log pages without rerunning the command. Shell-capable agents can also use the CLI with --json.

The agent can inspect the saved failure before checking the patch.
Recheck an existing unit test
Follow an exact NUnit or Unity test through a fix. Each attempt gets a fresh NUnit 3 report. Missing or skipped tests and unrelated failures stay inconclusive, so an agent cannot accept them as evidence that the selected test passed.
Connect your test or try the original EditMode example →
NUnit support is included in 1.3.0 through CLI and MCP. The documented Unity validation covers the Windows EditMode example.

The selected test stays the same; a skipped report is not accepted as a passing test.
Use it on your own failure
From your project, replace the command and message with your own:
npx --yes failtrace@1.5.0 run "npm test -- checkout" --repeat 20 --stderr-contains "checkout failed" --capture-contextRun this before editing in a Git project. --capture-context records source identity for Verify; outside Git, select files with --context-source. Each trial saves its output, and exit 1 can mean the target was captured successfully. Then edit and verify the patch →
Your next question | Command |
How often does this failure appear? |
|
What differs between a healthy and failing trial? |
|
Which Git change introduced it? |
|
What input is enough to reproduce it? |
|
What happened after the proposed fix? |
|
How can I replay this investigation? |
|
Command reference · Literal executable arguments · Reusable project scripts

The demo's bundle reproduces its original failure with exit 1. Keep the source, input and replay together; supply the target's prerequisites when packaging your own investigation.
What the results establish
FailTrace reports observations under the chosen settings. Verify separates a target observed, a healthy sample without that target, and inconclusive evidence. Bisect reports a sampled first-parent boundary; minimization rechecks its result without promising the smallest possible input.
The demo shows controlled example outcomes, not performance measurements. A passing sample does not prove a bug is gone.
Commands run with your permissions, and process cleanup is best effort. Retained stdout/stderr is capped by default at 16 MiB per trial and 256 MiB per run or bisect/minimization. Previous investigations accumulate separately. Review logs, commands and selected files before sharing; bundles still require the target's dependencies and setup.
Result and exit-code reference · Resource limits · Storage inventory · Bundle guide
Availability and contributing
The quick start uses npm's latest release; the verified version is 1.5.0. MCP configuration and repeatable installation examples keep that exact version pinned: installation options, GitHub release, and MCP Registry entry.
Version 1.5.0 adds: short run references, Verify readiness and next-step guidance, and intermittent-minimization guidance. See the changelog.
Documentation: choose your next task →
Development instructions · Compatibility · Roadmap · Third-party notices
This server cannot be deployed
Maintenance
Related MCP Connectors
QA platform for agents: coverage signals, in-repo test plans, verified tests and release governance.
Run, debug and inspect Playwright E2E tests from any AI agent: diagnostics, live DOM, selectors.
Recall your team's coding-agent memory. Install the Assertion plugin to capture it automatically.
- vibsyncOAuthcom.vibsync
One shared brain for your AI coding agents: team memory, agent Q&A, tasks, and file claims.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables coding agents to access recorded browser flows (user actions, network, console, etc.) for debugging and regression testing without reproducing issues.110Apache 2.0
- AlicenseNot gradedqualityBmaintenanceLocal-first MCP and coding-agent reliability harness that captures bounded, sanitized failure evidence and generates deterministic executable regression tests. Capture is opt-in; no API key or hosted service is required.2Apache 2.0
- AlicenseAqualityAmaintenanceEnables AI coding agents to retrieve relevant execution context before a task, compare candidate paths, log outcomes, and search past experiences stored locally in SQLite. It lets agents reuse what worked on similar tasks and get live alerts when a run stagnates or repeats failures.414200 PyPI2MIT
- AlicenseNot gradedqualityBmaintenanceEnables AI coding agents to record and resume engineering work as durable, byte-exact threads of focus, decisions, evidence, and next steps.2Functional Source , Version 1.1, ALv2 Future