Skip to main content
Glama
stopthegears

Token Compressor MCP

by stopthegears

Token Compressor MCP

Remote MCP server for Claude / Claude Cowork that compresses long context, stores compact reusable summaries, retrieves relevant context by title/tag/query, estimates token savings, and deletes stored records.

Tools

compress_context — compress pasted text without saving it
save_context — save raw text and a compressed summary
retrieve_context — retrieve saved compressed context by query, tag, title, or ID
list_contexts — list saved records
context_stats — show estimated savings
delete_context — delete one record
clear_contexts — delete all records

Related MCP server: Memory Cache MCP Server

Local setup

npm install
npm run dev

Local endpoints after start:

• MCP endpoint: http://localhost:3000/mcp
• Inspector: http://localhost:3000/inspector

Manufact deployment

Manufact / mcp-use projects support npm run dev, npm run build, npm run start, and npm run deploy. The mcp-use docs say local servers expose /mcp for MCP clients and /inspector for testing.

Recommended cloud settings:

Node version: 20+
Build command: npm install && npm run build
Start command: npm run start
MCP endpoint: https://<your-manufact-url>/mcp
Inspector: https://<your-manufact-url>/inspector

Claude Cowork usage

After deployment, add the remote MCP URL as a Claude custom connector.

Test prompts:

Use the token compressor MCP to compress this context without saving it:
[paste non-sensitive text]
Use the token compressor MCP to save this as "Test Context" with tags test, sample:
[paste non-sensitive text]
Retrieve the compressed context for Test Context.

Privacy note

This starter uses in-memory storage. Saved records disappear when the process restarts. That is intentional for a safer first version. Do not store confidential company data in a personal/free cloud server unless approved.

Related MCP Connectors

Related MCP Servers

  • F
    license
    B
    quality
    D
    maintenance
    A Model Context Protocol (MCP) server that optimizes token usage by caching data during language model interactions, compatible with any language model and MCP client.
    4
    2
    -
  • F
    license
    B
    quality
    D
    maintenance
    A Model Context Protocol server that reduces token consumption by efficiently caching data between language model interactions, automatically storing and retrieving information to minimize redundant token usage.
    4
    25
    -
  • A
    license
    C
    quality
    A
    maintenance
    An MCP server that filters and compresses context by 80-90% before sending to an LLM, using code knowledge graphs and compression.
    20
    2
    MIT