"A guide to understanding AI agents" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Simulate, test, and analyze cloud architectures without deploying real infrastructure. Cloud World Model enables AI agents to model cloud environments, evaluate architecture behavior and costs, run failure and chaos simulations, and explore infrastructure scenarios across cloud providers.
Check that your AI is being logical. Free tool that mathematically catches contradictions in agent reasoning. No account needed. Also offers paid guardrails that converts natural language to formal verification proofs, that anyone can check succinctly.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Drive real Android & iOS devices and web browsers from natural language for mobile + web QA. 290+ tools across device control, app management, automation sessions, browser automation, and flow recording / replay. Bearer-auth — get a token at robotactions.com → Profile → API Tokens.
Free platform to test MCP clients without installing anything. Create mock tools with dynamic templates, configurable delays, conditions (if/then), and response sequences. Supports JSON-RPC 2.0 over Streamable HTTP. Built-in text_echo and json_echo tools. Rate-limited tiers: anonymous (5 calls/min, 1 mock tool), registered (10 calls/min, 4 mock tools), premium (60 calls/min, unlimited). Zero setup — no install, no registration required. More info: https://www.testmcp.dev
Testing, benchmarking and auditing autonomous AI agents — methods, harnesses, evidence
Give AI assistants the context behind client website feedback. Read comments, screenshots, replies, element details, and developer briefs; organize priorities, update statuses, and export feedback from authorized projects. Coding agents with repository access can investigate issues and prepare fixes for review. Connect through Streamable HTTP and browser OAuth using a Simple Commenter account.
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Kilotest leverages 10 rule engines to test web pages for front-end quality (accessibility, usability, and standards conformity) and report results with specifiable granularity.
A webhook inbox for agents: one call returns a live URL. Mock, verify, inspect and replay.
Author and validate Calaf workspace seeds against the app's real importer. No account needed.
Generate synthetic random user data for testing, demos, and development without using real persona.
Workflow planning, recovery checkpoints, coordination, fixtures, and compatibility tools for agents.
Runs your code against a contract; HELD or BROKE at the exact input. Deterministic. 0.10 USDC/call.
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.
Score any URL against a real design contract — 42 checks, A-F grade, token + motion validation.
Whether a registry MCP server works for a stock client, and what changed in its tools. Free, no key.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Lint a SKILL.md for frontmatter, structure, secrets and size. All 6 tools free.