"A guide to CI/CD (Continuous Integration and Continuous Deployment)" matching MCP connectors:
Matching Connector Tools:
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Test-inbox API for email and SMS: create inboxes, long-poll messages, extract OTPs and links.
Free platform to test MCP clients without installing anything. Create mock tools with dynamic templates, configurable delays, conditions (if/then), and response sequences. Supports JSON-RPC 2.0 over Streamable HTTP. Built-in text_echo and json_echo tools. Rate-limited tiers: anonymous (5 calls/min, 1 mock tool), registered (10 calls/min, 4 mock tools), premium (60 calls/min, unlimited). Zero setup — no install, no registration required. More info: https://www.testmcp.dev
Check that your AI is being logical. Free tool that mathematically catches contradictions in agent reasoning. No account needed. Also offers paid guardrails that converts natural language to formal verification proofs, that anyone can check succinctly.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Promotion gate for AI agents: leakage audits, exact-statistics verdicts, and a live report card.
Hire a real human for real-world verification, product testing, AI output review, and errands.
Check AI work against requirements and return structured verdicts, findings, and repair steps.
Query Checkly synthetic monitoring — checks, statuses, results, alerts, reporting and dashboards.
Voice-powered bug reporting with 13 MCP tools. Record bugs by talking; let AI find and fix them.
One-cent x402 and MCP readiness checks plus fixed-price marketplace launch packs.
Generate deterministic placeholder image URLs and packs for docs, staging, testing, and AI agents.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
Scan URLs for WCAG 2.1 violations, generate AI fixes, and produce VPAT 2.5 compliance reports.
MCP-native AI evaluation: rubric audits, eval suites, and proof reports for AI/LLM output.
Read-only verifier for 25 ProofRelay MCP tools and non-confidential evidence bundles.
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.