HexWitness
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@HexWitnessShow me evidence for the message length validation"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Yesterday, your agent found the parser, mapped its callers, and proved which runtime event reached it. Today, a new chat opens and asks the live disassembler to discover everything again.
HexWitness stops that loop.
flowchart LR
V["Binary Ninja · IDA · Ghidra"] --> X["Reviewed evidence"]
R["Frida · debugger · wire observer"] --> C["Sealed capture"]
C --> X
X --> E[("HexWitness memory")]
E --> A["Your AI agent"]
A --> Q["Answer with proof"]
A -.->|one named gap| VThe live viewer remains the agent's eyes. HexWitness becomes its case file: tied to an exact build, searchable across tools, honest about conflicts, and ready for the next agent.
See it work in 60 seconds
Requirements: Git and Node.js 22.13 or newer.
npm install --global hexwitness
hexwitness setup
hexwitness demoThe setup wizard connects HexWitness to Codex, Claude Code, Cursor, VS Code/Copilot, Claude Desktop, or generic MCP clients. It also installs guidance written for that agent. The MCP entry starts the local read-only daemon when needed, so there is no service choreography.
Now ask:
Use HexWitness to explain 0x401120 in build toy-v1.
Show proof, contradictions, and missing evidence separately.The demo uses synthetic, redistributable evidence. No third-party binary data ships with HexWitness.
Prefer a checkout?
git clone https://github.com/siaginw/HexWitness.git
cd HexWitness
npm install
npm run demo
npm run setup -- --client codex --viewer none --yes
npm run doctorReplace codex with claude-code, cursor, copilot, or generic as needed. The setup command resolves the checkout's absolute local runtime path, installs guidance tailored to the selected agent, and writes the MCP connection without requiring a global install. Restart the selected AI client after setup, then ask it to call hexwitness_health and hexwitness_contract. Both must succeed before using private evidence.
The checked-in.mcp.json.example assumes npm install --global hexwitness. From a source checkout, run npm run setup instead of copying that example; setup records the correct local Node.js and bundled-runtime paths.
HexWitness installs as one command. The service, MCP transport, installer, capture pipeline, and adapter catalog live behind that command:
hexwitness agent # daemon autostart + MCP for AI clients
hexwitness serve # REST daemon only
hexwitness adapters # list every included viewer/runtime adapter
hexwitness adapters binary-ninja # print one adapter's exact path and capabilities
hexwitness contract # inspect the stable 1.x public contract
hexwitness backup ./evidence.db # create and verify a consistent snapshotOfficial MCP Registry identity: io.github.siaginw/hexwitness. Registry-aware clients discover the same local stdio server published through npm; its declared launch contract is hexwitness agent.
The npm package ships one bundled runtime instead of exposing its internal module tree. Python remains only in the thin Binary Ninja, IDA, and Ghidra exporters because those products expose their supported automation APIs through Python. Large JSONL exports and long captures are streamed through atomic ingest and disk-backed normalization instead of being loaded wholesale into memory.
Related MCP server: memory-bank-mcp
Why HexWitness feels different
Most RE integrations solve access: let an agent read a decompiler, debugger, or trace. HexWitness solves continuity: keep the useful result after that tool, build, or chat is gone.
Approach | Excellent for | The gap HexWitness fills |
Viewer MCP | Live decompilation, xrefs, renames, analysis control | Findings are session-scoped unless promoted |
Notes and reports | Human narrative | Hard to query, compare, or trace back to exact evidence |
General agent memory | Preferences and broad project context | No dedicated build, address, call-graph, capture, or provenance contract |
HexWitness | Durable RE evidence and runtime reconstruction | Pairs with viewers instead of replacing them |
That makes HexWitness especially useful when:
the binary changes and addresses move;
several agents or analysts share work;
static code must line up with runtime behavior;
a conclusion needs a reproducible chain of evidence;
a failed capture must be compared with a working one;
“we think” needs to become “we proved.”
Read the honest category comparison in Why HexWitness.
Ask real questions
Users ask about the target. The agent chooses the tools.
Which function validates frame length before dispatch?
Reuse retained evidence first. Inspect a live viewer only if one exact edge is missing.Compare the working and failing login captures.
Find the first meaningful divergence, then resolve its static consumer.Where is this UUID used, which class owns it, and did its field offset change
between builds?HexWitness gives agents first-class queries for builds, functions, classes, UUIDs, types, fields, vtables, calls, xrefs, paths, dataflow, captures, contradictions, coverage, evidence gaps, durable investigations, failed attempts, challenges, and discovery-only retrieval. Four MCP prompts package investigation, runtime comparison, live-finding promotion, and adversarial evidence-review workflows.
One investigation loop
Remember. Query retained evidence before touching a live tool.
Pin. Select the exact artifact build; never carry an address across builds by habit.
Resolve. Search the subject, then read its full evidence dossier.
Narrow. Traverse only the calls, fields, slices, or capture window needed.
Challenge. Surface provenance, confidence, and contradictory claims.
Escalate. If proof ends, name the smallest missing live observation.
Promote. Export that bounded result so the next investigation starts smarter.
For longer work, deterministic playbooks seed persistent checklists and operation budgets. Failed methods stay searchable. An evidence challenge surfaces opposition, unsupported claims, and open gaps without allowing agent consensus to inflate confidence.
Agents can also run allowlisted local RE utilities through one explicit MCP tool: argv-only, cwd-rooted, timed, output-capped, and receipt-producing. It is not an OS sandbox. Tool output remains an observation until promoted through the evidence model. No environment enable switch or separate model-provider key is required. See Investigation workbench.
The database remembers evidence. A separate privacy-safe activity store remembers operation hashes, timing, status, and counts—never prompts, arguments, or returned evidence.
Runtime capture without command soup
Collectors place conventional files and a small manifest in one private folder:
roundtrip/
├── capture.json
├── wire.jsonl
├── hooks.jsonl
├── screen.mp4
└── context.jsonThen:
hexwitness capture ./private/roundtripHexWitness checks the required roles and action markers, copies artifacts into an isolated private pack, normalizes a safe timeline by removing secret fields and replacing payloads with length plus SHA-256, seals checksums, verifies integrity, and imports the evidence. Missing or empty baseline artifacts fail closed. A failed normalization leaves the source pack recoverable.
See Capture packs for collector and scenario contracts.
Works with your stack
Tool | Durable bridge |
Binary Ninja | Deep JSONL exporter plus optional official Binary Ninja MCP live viewer |
IDA / IDAPython | Functions, strings, imports, references, blocks, and optional Hex-Rays pseudocode plus live MCP |
Ghidra | Functions, strings, imports, references, blocks, types, fields, and enum exporter |
Frida 17 | Narrow semantic-event observer and fail-closed JSONL normalizer |
Other tools | Versioned adapter manifest and vendor-neutral JSONL schema |
HexWitness does not ship a weaker disassembler inside the project. Viewer MCPs provide live eyes. Exporters turn reviewed findings into portable memory. Read Viewer MCP bridges and the Adapter SDK.
One truth, three interfaces
Interface | Best use |
MCP | Agent-led investigations, read-only evidence tools, and one explicitly open-world local-tool runner |
CLI | Import, capture lifecycle, automation, recovery, and direct queries |
REST | Local read-only integration with a published route manifest |
All three read the same SQLite evidence graph. Imports are transactional and idempotent. Addresses remain canonical hexadecimal strings, including unsigned 64-bit values.
Evidence, not vibes
Build — exact artifact identity and analysis provenance.
Entity — function, type, class, field, vtable, runtime object, or project-defined object.
Edge — call, reference, ownership, inheritance, dataflow, or runtime relationship.
Evidence — an observation with source, time, classification, and confidence.
Claim — an interpretation linked to supporting or opposing evidence.
Capture — a sealed scenario with artifacts, markers, events, and checksums.
Gap — the next missing fact needed to prove a behavior.
Conflicting claims remain visible. Unknown behavior becomes a worklist, not a fabricated answer.
Private by default
raw private material → normalized project evidence → synthetic public fixturesDaemon binds to localhost and serves GET-only queries, including a loopback-only dashboard.
Non-local binding requires an API token and still needs trusted TLS transport.
Executable bytes and decompiler text are not exported by default.
Capture normalization recursively removes common secret and payload fields.
Public-release audit blocks credentials, captures, dumps, proprietary binary formats, large embedded payloads, and personal paths.
Read Privacy and Security before importing sensitive work.
Documentation
Start here | What it answers |
Can I prove the full local loop? | |
Why not just use notes, a viewer MCP, or generic memory? | |
What gets installed for each agent? | |
What should I ask the agent? | |
What is complete, variable, or intentionally out of scope? | |
Which claims are machine-checked? | |
What remains compatible throughout 1.x? | |
Which runtimes and viewer boundaries are supported? | |
What does the 1.0 claim include? | |
What interfaces are available? | |
How do the pieces fit? | |
What failed and how do I prove it? |
Project status
HexWitness 1.0 is a stable public developer release. Automated gates cover schema migration, importer, evidence graph, query engine, capture lifecycle, unified CLI, bundled distribution, read-only daemon, concurrent query behavior, MCP server, setup wizard, tailored agent skills, privacy audit, packaging, installed upgrade, and the CLI → DB → daemon → MCP journey. The exact trust boundary is documented in Release readiness.
Commercial viewer APIs still vary by edition and release. HexWitness documents that boundary instead of claiming universal compatibility. See Quality for the exact tested surface.
Contributing
Focused issues and pull requests welcome. Read Contributing. Never attach proprietary binaries, vendor databases, credentials, or captures you cannot redistribute.
Apache-2.0. Analyzed binaries, imported evidence, vendor SDKs, and RE databases retain their own terms.
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Connectors
Persistent, inspectable memory for AI agents with lineage, correction, and a hosted MCP endpoint.
A paid remote MCP for agent memory MCP, built to return verdicts, receipts, usage logs, and audit-re
Shared memory for all your AI agents, your whole team and every MCP client — save, search, recall.
MCP-native Trust Infrastructure for AI Agents. Persistent encrypted memory with Trust Quotient.
Related MCP Servers
AlicenseNot gradedqualityAmaintenanceProvides AI agents with a live architecture model of a codebase, enabling queries for root cause analysis, blast radius, and dependency traversal through MCP tools.23818Apache 2.0- AlicenseBqualityBmaintenanceProvides a read-only MCP interface to query and retrieve verifiable evidence from a local memory bank, supporting search, dossier, chronology, source, and evidence tools.6BSD Zero Clause
- AlicenseNot gradedqualityDmaintenanceProvides coding agents with governed semantic memory and code-graph context via MCP, enabling code-linked recall, blast-radius impact analysis, and lifecycle-aware memory management.2Apache 2.0
- FlicenseAqualityCmaintenanceA read-only MCP server that gives AI coding agents structured access to a project's source code, architecture, documentation, and Git context through 16 tools for searching, reading, and comparing evidence without modifying files.16
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/siaginw/HexWitness'
If you have feedback or need assistance with the MCP directory API, please join our Discord server