mcp-agent-reliability-scorer
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@mcp-agent-reliability-scorerScore this agent trajectory for reliability issues"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
mcp-agent-reliability
Pure-computation MCP server that scores AI agent trajectories, detects silent failures, loops, and reliability metrics — zero external API cost.
What problem does this solve?
AI agents often fail silently: they loop on the same tool, return empty results, or give high-confidence answers without evidence. Debugging is expensive. This MCP gives you instant scores and failure reports from any agent run log you provide.
Think of it like a doctor’s check-up for your AI agent — it looks at the “X-ray” (the run log) and tells you what is healthy and what is broken, without needing another expensive doctor (LLM).
Related MCP server: HumanProof
Tools
Tool | Description |
| 0-100 reliability score + breakdown |
| List of loops, empty results, high-confidence-without-evidence, etc. |
| Per-tool call counts, error rates |
| Success rate across many runs |
| Which of two runs is more reliable |
Quick Start (local)
# Install
pip install -e .
# Run (HTTP on port 8080)
python -m mcp_agent_reliability.serverOr with MCP inspector / Claude Desktop / Cursor by pointing to the HTTP endpoint.
Example trajectory input
[
{"tool": "get_weather", "status": "ok", "result": {"temp": 28}},
{"role": "assistant", "content": "28C today", "is_final": true}
]Why this is valuable for entrepreneurs
Zero running cost (no paid APIs)
Helps you ship reliable agents faster → happier users → more revenue
Can be called by the agent itself mid-run or by your CI after tests
Fits the “Type A” high-margin MCP pattern preferred on MCPize
Development
pytestLicense
MIT
Built daily for Prince Ruhul / Prevalid by the Daily AI Project Builder.
This server cannot be deployed
Maintenance
Related MCP Connectors
AI agent observability for production traces, natural-language insights, and improvement loops.
Synthetic checks, nightly regression replay and model-drift alerts for AI agents
Deterministic runtime safety for AI agents: scan PII, gate tool actions, verify LLM output.
See, price, and control every tool call your AI agents make: policy checks, cost, and audit tools.
Related MCP Servers
- AlicenseNot gradedqualityAmaintenanceEnables AI agents to retain memory of past interactions and detect behavioral drift, preventing repeated mistakes without LLM token extraction.15 npm139 PyPI207MIT
- AlicenseNot gradedqualityAmaintenanceAgent trajectory scorer for human-likeness — flags bot-like patterns in automated workflows before they reach production.MIT
- AlicenseBqualityCmaintenanceValidates agent outputs in multi-agent systems to prevent coordination failures, with tools for schema verification, hallucination detection, and freshness checks, all with zero LLM cost.524 npmMIT

Pisama MCP Serverofficial
AlicenseNot gradedqualityAmaintenanceEnables analysis of AI agent traces to detect and fix failures using heuristic detectors, with no LLM calls required.225 PyPI1MIT