"Information about playwrights or the Playwright testing framework" matching MCP connectors:
Matching Connector Tools:
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
Hire a real human for real-world verification, product testing, AI output review, and errands.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Generate deterministic placeholder image URLs and packs for docs, staging, testing, and AI agents.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.
Read-only MCP server for the OPERANT AI operating-agent calibration benchmark.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
MCP-native AI browser testing for coding agents. Submit a URL + goal, get back action trail, bugs, screenshots, and WebM video your agent patches from directly. 43 tools, 12 AI evaluation personalities, combo tiers with auto-pause-on-bugs, throwaway email + SMS inboxes.
130+ QA & dev tools for AI agents: prompt injection, RAG testing, VLM eval, guardrails. Free.
MCP server for e-mail testing: create disposable inboxes, wait for delivery, and extract e-mail content or links - all from your AI agent or test automation workflow. Get a free API key on https://app.zyntra.app/
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.
End-to-end API testing — generate and run tests from OpenAPI, curl, Postman, or real user traffic.
ResilienceOracle - 10 operational resilience tools: BIA, RTO/RPO, scenario testing.
Drive OctoPerf load testing from any AI agent — import, edit, validate, run scenarios, read metrics. Hosted remote server, OAuth 2.1 (DCR + PKCE), no API key.
MCP server providing access to the Scorecard API to evaluate and optimize LLM systems.
Check AI work against requirements and return structured verdicts, findings, and repair steps.
Programmatic email deliverability testing for AI agents. Create inbox placement tests across Gmail, Outlook, Yahoo, Mail.ru, Yandex — get per-provider placement (Inbox/Spam/Promotions), SPF/DKIM/DMARC auth, Rspamd & SpamAssassin verdicts, DNS health, and live SSE results.