"A map or mapping service" matching MCP connectors:
Matching Connector Tools:
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Check that your AI is being logical. Free tool that mathematically catches contradictions in agent reasoning. No account needed. Also offers paid guardrails that converts natural language to formal verification proofs, that anyone can check succinctly.
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Hire a real human for real-world verification, product testing, AI output review, and errands.
Webhook relay for AI agents: keyless quickstart, event inspection, replay, mock payloads.
Promotion gate for AI agents: leakage audits, exact-statistics verdicts, and a live report card.
Email compatibility analysis across 15 clients — preview, audit, fix, diff, deliverability.
Query Checkly synthetic monitoring — checks, statuses, results, alerts, reporting and dashboards.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Read-only MCP server for the OPERANT AI operating-agent calibration benchmark.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
Scan URLs for WCAG 2.1 violations, generate AI fixes, and produce VPAT 2.5 compliance reports.
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
Read-only verifier for 25 ProofRelay MCP tools and non-confidential evidence bundles.
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
MCP-native AI browser testing for coding agents. Submit a URL + goal, get back action trail, bugs, screenshots, and WebM video your agent patches from directly. 43 tools, 12 AI evaluation personalities, combo tiers with auto-pause-on-bugs, throwaway email + SMS inboxes.