mcpgawk
OfficialClick on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@mcpgawkscan my MCP servers and block any tool that changed since I approved it"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
mcpgawk
Your agents call Model Context Protocol servers that can change what their tools do after you approved them, and the agent will call the new one without noticing. mcpgawk reads every server your agents can reach, checks every call against a baseline you approved, and blocks the ones that changed. It runs on your machine and uploads nothing.
The same engine powers mcpgawk Platform: mcpgawk enforce puts one endpoint in front of the
whole fleet with a key per caller, policy on every call and a hash-chained audit log, and
mcpgawk monitor watches the servers you approved around the clock and tells you when one drifts.
This free layer is the seeing and the blocking underneath it. mcpgawk Platform is one subscription
per person for up to 3 machines: start a free 7-day trial at https://mcp.gawk.dev/trial.html
(no card) or subscribe at https://mcp.gawk.dev/subscribe. Then mcpgawk login <key> on this same
install fetches the paid engine and turns those on โ one more command, nothing else to set up.
Why
A server you approved can change what its tools do afterwards. Nothing in MCP tells your agent that happened โ it just calls the new tool. That is the rug-pull, and it is the case mcpgawk is built for.
Two things follow from being able to see a server properly. You find out what each one can reach before you trust it, and you find out what it costs: every tool is loaded into your context on every request, used or not.
Related MCP server: Sentinel
How it's different
It blocks, it does not only report. A scanner tells you afterwards.
mcpgawk guardinstalls one pre-execution hook and a tool that appeared after you approved the server does not run.It runs the server, not just reads it.
mcpgawk verifydrives tools in a sandbox and reports what they actually did โ exfiltration, SSRF, poisoning โ reproduced before it is reported.Nothing is uploaded. Cloud scanners send your inventory to a server and gate the verdict there. Every decision here is made on your machine, with no account and nothing to sign in to.
It says what it did not check. Skipped tools are named as skipped, never counted as clean.
Features
๐ Block a changed tool before it runs โ
mcpgawk guard installputs one pre-execution hook in your agent's loop. The decision is local, in about 10ms, with nothing to sign in to. Works on 6 of the 21 supported clients; the rest have no hook point and are named, not glossed over.๐งช Run it, don't just read it โ
mcpgawk verifydrives tools in a sandbox and reports what they did: exfiltration, SSRF, tool poisoning, secret leaks. The sandbox is a proxy by default, so it needs no Docker; Docker adds full container isolation when you have it. Safe mode drives only provably read-only tools, and every tool it skips is named as skipped.๐งโโ๏ธ Approval needs a person โ
mcpgawk decideopens a local screen for what changed. The buttons live on the tokened link printed in your terminal, so an agent that opened the page cannot approve its own way past a block.๐ฅ๏ธ One local panel โ
mcpgawk panel: every server, every decision, every piece of evidence.๐ The changelog no vendor publishes โ
mcpgawk changesshows every change to a server's tool surface between the snapshots you have recorded: tools added or removed, input schemas widened, descriptions and annotations rewritten. It reads your local history, so it works for a server you approved weeks ago โ the one thing a fresh scan can never tell you.๐ Any transport โ stdio, streamable-HTTP, SSE, and OAuth remotes (via the
mcp-remotebridge).๐ธ Token cost index โ exactly what each tool adds to your context at connect, plus the 3 heaviest tools.
๐งพ Capability facts โ write / exfil-capable / declared annotations, straight from the schema, plus a trust-surface summary (% write, % exfil-capable, destructive-declared count) and an annotation-completeness score.
๐ Integrity pin + drift โ catch a server that silently rewrites its tools (
--track).๐ฉ Bounded signals โ injection-shaped descriptions, cross-server shadowing, under-declaring Server Cards โ pointers for a human, never verdicts.
๐ Zero egress, by construction โ the measurement layers import no network library. Enforced by a test. Two checks are opt-in and make an explicit exception (see Guarantees):
--supply-chainand--oauth-scopes.
Get it โ three ways
CLI (any terminal):
uv tool install --force mcpgawk # or: pipx install --force mcpgawk
mcpgawk # finds every agent config on the machine itselfRun it as an MCP server: mcpgawk mcp (stdio), or uvx mcpgawk mcp.
Editor (VS Code / Cursor): install mcpgawk from the marketplace (Open VSX). It scans your workspace mcp.json and shows cost + capability flags inline. The extension drives this engine as a subprocess โ it is built and released separately, so its source is not in this repository.
CI (GitHub Action): gate every PR on token budget / drift (Marketplace):
- uses: gawk-dev/mcpgawk@v1
with: { config: mcp.json, max-tokens: 8000, fail-on-flagged: true }When to run it
Once, then leave it on โ
mcpgawk guard install. After that a tool that appears on a server you already approved does not get called.Before you add a server โ see what it costs and what it can do, before you trust it.
When your agent feels slow or picks the wrong tool โ it's often MCP bloat (too many / too-heavy tools).
On every PR โ the CI gate catches drift and creeping token cost.
If you publish an MCP server โ see what it costs your users and how it reads to a client, and fix it (usually one line per tool). Lean + well-annotated is a differentiator.
Use
mcpgawk # first run: every agent config on this machine
mcpgawk demo # the whole arc in a sandbox โ approve, drift, block
mcpgawk guard install # put the baseline in your agent's loop
mcpgawk guard status # is protection actually on?
mcpgawk decide # what changed, and approve it as a human
mcpgawk panel # the local page: servers, decisions, evidence
mcpgawk verify mcp.json # run the servers and watch what they do
mcpgawk changes # what changed on your servers since you approved them
mcpgawk login <key> # licence? one command fetches the paid engine (enforce, monitor)Scanning on its own, if that is all you want:
mcpgawk scan mcp.json # a whole config
mcpgawk scan --stdio "npx -y @modelcontextprotocol/server-filesystem@2026.8.31 /tmp"
mcpgawk scan --http https://host/mcp --header "Authorization: Bearer $TOKEN"
mcpgawk scan --sse https://host/sse
mcpgawk scan mcp.json --track # record + detect rug-pulls over time
mcpgawk scan mcp.json --json # machine-readable labels
mcpgawk scan mcp.json --verbose # full per-tool table, not just flagged tools
mcpgawk scan mcp.json --supply-chain # opt-in: npm/PyPI deprecation check (network)
mcpgawk scan mcp.json --oauth-scopes # opt-in: decode a supplied Bearer JWT's scopeWhat it reports
Cost index โ tokens each tool adds at connect (named tokenizer; a comparable index, not an absolute Claude count), plus the 3 heaviest tools.
Trust surface โ capability facts (write/mutating, exfil-capable, declared annotations) rolled up into % write, % exfil-capable, and a destructive-declared count.
Annotation completeness โ a transparent composite (annotated รท total tools declaring read/write intent), not a risk score.
Coverage โ tools, prompts, and resources counted (
--verbosefor the full per-tool table).Integrity pin โ a hash that changes if the server silently rewrites its tools;
--trackturns it into rug-pull detection over time.Bounded signals โ precise, low-false-positive pointers for a human to review, never verdicts: injection-shaped descriptions (tools and prompts), cross-server name shadowing, and public Server Cards that under-declare what the server actually exposes.
Supply-chain (opt-in,
--supply-chain) โ checks the launched package against the public npm/PyPI registry for deprecation/yank status.OAuth scopes (opt-in,
--oauth-scopes) โ locally decodes a supplied Bearer JWT'sscopeclaim.
Guarantees
No inventory egress. The only network is the protocol client talking to the server you point it at. The measurement layers import no network library โ they cannot egress by construction (enforced by a test). Public Server Card discovery is fetched with no auth and no redirect-following. Two flags are the explicit, opt-in exception:
--supply-chainsends the launched package's name (and pinned version, if any) โ never your tool inventory โ to the public npm registry or PyPI JSON API.--oauth-scopesmakes no network call at all; it locally decodes a Bearer JWT you already supplied. Neither runs unless you pass the flag.Facts โ heuristics. Exact capability facts and the token index never mix with the bounded heuristic signals โ separate in code, separate in output.
Reproducible. One command, identical numbers.
Tracks the protocol. Built on the official
mcpSDK, which negotiates the protocol version.
Develop
uv run --extra dev --with mcp --with tiktoken --with httpx python -m pytest -qCI gate โ GitHub Action
Scan your MCP servers on every pull request and fail the build if one gets too heavy or trips a signal. It runs entirely in your runner โ nothing is uploaded โ and posts a per-server cost/flag table to the job summary.
- uses: gawk-dev/mcpgawk@v1
with:
config: mcp.json # or: stdio / http / sse โ a single server
max-tokens: 8000 # fail if any server loads more than this at connect
fail-on-flagged: true # fail if any bounded signal firesAvailable on the GitHub Marketplace.
Contributing
Issues and PRs welcome. Please read CONTRIBUTING.md first, and see the design boundaries in THREAT-MODEL.md. Security reports go through SECURITY.md (privately, not a public issue).
License
Apache-2.0 โ see LICENSE. Part of the nativerse ยท gawk.dev family. Site and docs: mcp.gawk.dev. The value is in the repo, not a cloud.
Use it from your agent (skill)
Let your coding agent run the checks itself โ whenever it adds, upgrades or audits an MCP server:
# Claude Code (similar for other agents: copy the folder into their skills directory)
mkdir -p ~/.claude/skills && cp -r skills/mcpgawk ~/.claude/skills/mcpgawkThe skill teaches the agent to measure a server BEFORE trusting it, audit an MCP-2 upgrade as a baseline diff instead of blind re-trust, and relay every consent prompt to you verbatim.
This server cannot be deployed
Maintenance
Related MCP Connectors
The MCP server that vets MCP servers: identity, risk grade and per-tool risk before you install.
Search, vet & assemble MCP servers from your agent: verified tools, risk labels, and trust scores.
- gatewayOAuthai.sealgate
MCP gateway with runtime security policy, tool-call-level control, and audit of agent actions.
Trust checks for MCP servers: trust scores, tool-drift detection, signed diligence receipts. Free.
Related MCP Servers
- AlicenseBqualityDmaintenanceA security-focused Model Context Protocol server that enables controlled local tool execution through strict network firewalls, filesystem protections, and rate-limiting policies. It features a plugin-based architecture for progressive tool discovery and includes reference implementations for web searching and bug tracking.15MIT
- AlicenseNot gradedqualityCmaintenanceA safe, allowlisted MCP server that lets AI agents run only a tiny set of harmless tools (echo, datetime, hash, dig, GET-only curl, whois, status checks) against explicitly allowed hosts, with sanitization, rate limiting, timeouts, and full audit logging.MIT
- AlicenseNot gradedqualityBmaintenanceMCP server that detects and guards against tool poisoning and prompt injection attacks in tool descriptions and schemas. It provides risk scoring, pattern detection, safe rewriting, and audit reports with zero external API cost.MIT
- FlicenseNot gradedqualityCmaintenanceMCP server that provides a security gateway for AI agents, enforcing allow/confirm/deny policies on tool calls and requiring human approval for risky operations, with full audit logging.-