"Understanding the Concept of Perplexity" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Diagnose why an AI agent failed and get the verified fix instantly. Free, no token.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Give AI assistants the context behind client website feedback. Read comments, screenshots, replies, element details, and developer briefs; organize priorities, update statuses, and export feedback from authorized projects. Coding agents with repository access can investigate issues and prepare fixes for review. Connect through Streamable HTTP and browser OAuth using a Simple Commenter account.
Author and validate Calaf workspace seeds against the app's real importer. No account needed.
Risk-scan a diff, flag AI-generated-code tells, find secrets. 5 of 7 tools need no account.
Runs your code against a contract; HELD or BROKE at the exact input. Deterministic. 0.10 USDC/call.
What is known to be broken in an MCP server or API operation, with the check that found it.
A deterministic verification engine for agents. Proves a fix: fails on the old code, passes on new.
Seven tools over the tabnas parsing engine: parse, validate, diagnose, fixtures, compare.
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
Ensemble testing of web pages for accessibility, usability, and standards conformity
Compare two versions of a JSON row list: what was added, removed or changed, field by field.
Checks the structural integrity of translated resource dictionaries against a source dictionary y...
Defectbird: the site's own MCP server — dataset; every answer cites the site.
Tests an AI agent's purchase against the task it was given. Paid per call in USDC via x402.
Class Q Checker: the site's own MCP server — checker, enquiry (enquiry = a human handoff, not a...
Smoke Control Checker: the site's own MCP server — checker, enquiry (enquiry = a human handoff,...
Evidence-bound second-opinion audit of an agent conclusion against caller-supplied evidence.