genpark-agent-semantic-cache-similarity-deduplicator-skill
OfficialClick on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@genpark-agent-semantic-cache-similarity-deduplicator-skillCheck the semantic cache for similar queries before answering: what is MCP?"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
genpark-agent-semantic-cache-similarity-deduplicator-skill
โก Overview & Architectural Significance
genpark-agent-semantic-cache-similarity-deduplicator-skill delivers zero-dependency, low-latency, deterministic agentic memory and retrieval primitives engineered strictly using Python 3.9+ standard library.
๐ Key Architectural Capabilities
Zero External Dependencies: Operates exclusively via pure Python (
math,re,collections,heapq,hashlib,json). Zero pip install overhead, zero C-extension compile errors.Enterprise RAG & Memory Invariants: Implements formal algorithms for cognitive decay, BM25 Okapi lexical scoring, Reciprocal Rank Fusion, knowledge graph traversal, semantic query caching, and lost-in-the-middle context reordering.
Native Anthropic MCP Protocol: Compliant with standard JSON-RPC 2.0 stdio MCP specifications for Claude Desktop, Cursor, and Windsurf.
Related MCP server: Arda Vector Database MCP Server
๐๏ธ Architectural Topology & State Machine
flowchart TD
UserQuery["User Prompt / Agent Goal"] --> SemCache["Semantic Cache Check"]
SemCache -->|Cache Hit| FastReturn["Cached Response (0ms)"]
SemCache -->|Cache Miss| DualRetrieval["Dual Retrieval Pipeline"]
subgraph DualRetrieval ["Hybrid Search Engine"]
BM25Lex["BM25 Okapi Lexical Ranker"]
DenseVec["Dense Cosine Vector Similarity"]
end
DualRetrieval --> RRF["Reciprocal Rank Fusion (RRF)"]
RRF --> GraphExp["Knowledge Graph Triplet Expansion"]
GraphExp --> LostMiddle["Lost-In-The-Middle Context Reorderer"]
LostMiddle --> LLM["LLM Synthesis with Optimal Context"]
LLM --> EpisodicMem["Episodic Consolidation & Recency Decay"]๐ Quickstart & Standalone Execution
Local Python Client Usage
from client import AgentSemanticCacheDeduplicator
# Initialize engine
engine = AgentSemanticCacheDeduplicator()
# Execute self-testing benchmark suite
result = engine.run_benchmark_semantic_cache()
print("Execution Result:", result)๐ One-Click MCP Integration (Claude Desktop / Cursor)
Add to your claude_desktop_config.json or cursor.json:
{
"mcpServers": {
"genpark-agent-semantic-cache-similarity-deduplicator-skill": {
"command": "python",
"args": ["-u", "/path/to/genpark-agent-semantic-cache-similarity-deduplicator-skill/mcp_server.py"]
}
}
}๐ฆ Smithery.ai & PyPI Deployment
This skill contains pre-configured smithery.yaml and pyproject.toml manifests. Install directly via pip:
pip install git+https://github.com/alphaparkinc/genpark-agent-semantic-cache-similarity-deduplicator-skill.gitThis server cannot be deployed
Maintenance
Related MCP Connectors
Persistent semantic memory storage, associative recall, and recent memory index by namespace.
Shared knowledge cache for AI coding agents โ reuse an answer once it exists.
LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.
Hosted persistent memory with semantic search, importance and TTL for AI agents.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables storing and retrieving information using semantic search with Qdrant vector database. Acts as a memory layer for LLMs to persistently store and semantically search through information and metadata.Apache 2.0
- FlicenseNot gradedqualityDmaintenanceEnables semantic code search across multi-language codebases using natural language queries, integrated with Qdrant vector database for fast, cached retrieval.1-
- -licenseNot gradedqualityNot gradedmaintenanceSemantic caching MCP server for AI agent tool calls, providing exact and similarity-based cache lookup, store, invalidation, and metrics via MCP tools.-
- FlicenseNot gradedqualityBmaintenanceEnables agents to deduplicate semantic prompt caches by measuring Jaccard token overlap, preventing redundant LLM calls and reducing cost. It provides deterministic lookup-or-insert cache behavior through an MCP-compatible interface.8-