"Read the Docs" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Author and validate Calaf workspace seeds against the app's real importer. No account needed.
Moderated usability testing: read sessions, notes, transcripts, reports, and draft test scenarios.
Checks the structural integrity of translated resource dictionaries against a source dictionary y...
Tests an AI agent's purchase against the task it was given. Paid per call in USDC via x402.
Seven tools over the tabnas parsing engine: parse, validate, diagnose, fixtures, compare.
Point Claude Code, Qwen Code, Cursor, or any MCP client at https://docs.jmeter.ai/api/mcp and your agent answers JMeter questions grounded in this documentation, with a source link for every answer. Free, no API key, no signup.
Benchmark-first release surface with a read-only MCP endpoint and operator CLI.
You are the model under test. Enter ScoreIA Open Chamber; signed cards include failures. Auth none.
Read-only, deterministic AI triage and readiness tools implementing Sophon's published rubrics.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Deterministic recipe verification engine — validates AI-generated recipes against master SOPs.
Generate deterministic placeholder image URLs and packs for docs, staging, testing, and AI agents.
Read-only MCP server for the OPERANT AI operating-agent calibration benchmark.
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.