"Information about playwrights or the Playwright framework" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Author and validate Calaf workspace seeds against the app's real importer. No account needed.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Checks the structural integrity of translated resource dictionaries against a source dictionary y...
Check what a paid x402 endpoint or MCP server delivered, from probes anyone can repeat.
Defectbird: the site's own MCP server — dataset; every answer cites the site.
Class Q Checker: the site's own MCP server — checker, enquiry (enquiry = a human handoff, not a...
Smoke Control Checker: the site's own MCP server — checker, enquiry (enquiry = a human handoff,...
Tests an AI agent's purchase against the task it was given. Paid per call in USDC via x402.
Deterministic operations reconciliation for AI agents: COMPLETE, INCOMPLETE, or NEEDS_REVIEW.
Seven tools over the tabnas parsing engine: parse, validate, diagnose, fixtures, compare.
Point Claude Code, Qwen Code, Cursor, or any MCP client at https://docs.jmeter.ai/api/mcp and your agent answers JMeter questions grounded in this documentation, with a source link for every answer. Free, no API key, no signup.
Scan any website or MCP server for agent readiness: 0-100 score, a fix per failing check. Free.
You are the model under test. Enter ScoreIA Open Chamber; signed cards include failures. Auth none.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Read-only MCP server for the OPERANT AI operating-agent calibration benchmark.
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes