"Researching an author based on their books" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Verify before your agent acts on data it paid for. Signed verdicts, checkable offline, via x402.
Author and validate Calaf workspace seeds against the app's real importer. No account needed.
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Test a Python package/version on a target runtime. Scoped Linux install and usable-import evidence.
Probe a signup URL you own and score whether an AI agent can sign up unaided.
Scores any public website on how usable it is by AI agents, with per-check evidence.
Resolve whether an uncertain side-effecting action completed before software retries it.
Tests an AI agent's purchase against the task it was given. Paid per call in USDC via x402.
Verify an agent's advertised route, price, payment details, and schemas against its live endpoint.
MCP tool observatory: do registry servers answer, and are their answers true? No key.
AgentReady.market audit: can an AI shopping agent find, understand and BUY on this store? /100.
Rule-based site audits: accessibility, SEO, security headers, performance. Metered per call.
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.
Pay-per-call MCP server. Vetted human experts review AI-generated content (text, images, video, audio, social posts), audit reasoning chains, run prepublish safety checks, and gate high-stakes actions for human approval. 7 paid tools at $1.00 each plus 4 free tools (list_offerings, list_expert_profiles, get_result, verify_certificate). Payment via x402, USDC on Base mainnet. Approved outputs receive an on-chain Taste content certificate downstream agents verify before consuming.
MCP-native AI browser testing for coding agents. Submit a URL + goal, get back action trail, bugs, screenshots, and WebM video your agent patches from directly. 43 tools, 12 AI evaluation personalities, combo tiers with auto-pause-on-bugs, throwaway email + SMS inboxes.
Evaluates UI designs for WCAG accessibility issues automated scanners miss. Paid via x402 on Base.
Test-inbox API for email and SMS: create inboxes, long-poll messages, extract OTPs and links. 34 tools; bearer auth with an mfx_ API key (free tier).
remote debug iOS/Android/Unity/Godot/Flutter/RN/Web on real-device.ui-tree/screenshots/taps,tests.
Pluralistic human evaluation infrastructure for AI in production. Query real human reviewer verdicts on commercial AI models - scores, flag breakdowns, and side-by-side comparisons, by our community reviews.