"An overview of knowledge graphs" matching MCP connectors:
Matching Connector Tools:
Conformance checker for MCP servers. Free, no key, verdicts recomputable and re-measured daily.
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
A skeptical senior-engineer code reviewer over MCP: risk-scans unified diffs, flags AI-generated-code tells, reports complexity hotspots, scans for leaked secrets, and runs an OWASP security pass — real analyzers, no external APIs. Free tier, no signup.
AI integrity standards, benchmarks, and EU AI Act-aligned Tier 0 model certification
Resolve whether an uncertain side-effecting action completed before software retries it.
AgentReady.market audit: can an AI shopping agent find, understand and BUY on this store? /100.
Preflight QA for AI-agent deliverables with structured verdicts and repair guidance.
Hire a real human for real-world verification, product testing, AI output review, and errands.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
Scan URLs for WCAG 2.1 violations, generate AI fixes, and produce VPAT 2.5 compliance reports.
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
Pay-per-call MCP server. Vetted human experts review AI-generated content (text, images, video, audio, social posts), audit reasoning chains, run prepublish safety checks, and gate high-stakes actions for human approval. 7 paid tools at $1.00 each plus 4 free tools (list_offerings, list_expert_profiles, get_result, verify_certificate). Payment via x402, USDC on Base mainnet. Approved outputs receive an on-chain Taste content certificate downstream agents verify before consuming.
130+ QA & dev tools for AI agents: prompt injection, RAG testing, VLM eval, guardrails. Free.
MCP server for static security analysis of Android source code
Test-inbox API for email and SMS: create inboxes, long-poll messages, extract OTPs and links. 34 tools; bearer auth with an mfx_ API key (free tier).
MCP server providing access to the Scorecard API to evaluate and optimize LLM systems.
295k+ bug-fix patterns with MCP Hub proxy, PII filtering, and code search