"An overview of temporal knowledge graphs" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Diagnose why an AI agent failed and get the verified fix instantly. Free, no token.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Risk-scan a diff, flag AI-generated-code tells, find secrets. 5 of 7 tools need no account.
What is known to be broken in an MCP server or API operation, with the check that found it.
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Test an email before it goes out: a disposable address, 41 checks with RFC citations, a fix plan.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
Pay-per-call MCP server. Vetted human experts review AI-generated content (text, images, video, audio, social posts), audit reasoning chains, run prepublish safety checks, and gate high-stakes actions for human approval. 7 paid tools at $1.00 each plus 4 free tools (list_offerings, list_expert_profiles, get_result, verify_certificate). Payment via x402, USDC on Base mainnet. Approved outputs receive an on-chain Taste content certificate downstream agents verify before consuming.
Ensemble testing of web pages for accessibility, usability, and standards conformity
Probe a signup URL you own and score whether an AI agent can sign up unaided.
Compare two versions of a JSON row list: what was added, removed or changed, field by field.
Resolve whether an uncertain side-effecting action completed before software retries it.
Checks the structural integrity of translated resource dictionaries against a source dictionary y...
Tests an AI agent's purchase against the task it was given. Paid per call in USDC via x402.
Evidence-bound second-opinion audit of an agent conclusion against caller-supplied evidence.
AgentReady.market audit: can an AI shopping agent find, understand and BUY on this store? /100.
Verify an agent's advertised route, price, payment details, and schemas against its live endpoint.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.