"Cal.com" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
Accessibility and WCAG data for your own websites: fix lists, live checks, and fix validation.
Lint a SKILL.md for frontmatter, structure, secrets and size. All 6 tools free.
Risk-scan a diff, flag AI-generated-code tells, find secrets. 5 of 7 tools need no account.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Validate HTML/CSS, audit SEO and JSON-LD, check links, and capture responsive screenshots.
Ensemble testing of web pages for accessibility, usability, and standards conformity
Regex matching, email/URL format validation, and JSON Schema validation.
Kilotest leverages 10 rule engines to test web pages for front-end quality (accessibility, usability, and standards conformity) and report results with specifiable granularity.
MCP tool observatory: do registry servers answer, and are their answers true? No key.
Exposes FEDLIN's public security scanners as agent-callable tools over Streamable HTTP.
49 deterministic tools for text integrity, agent control, and contextual quality evidence.
MCP server: NEED + YIELD + CLEAN-MONEY gates with EIP-3009 attestations · Hive Civilization
Find MCP servers and check whether they actually respond, via live handshake probes.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Scores any MCP server before you trust it: free quick check, full paid report, 2-5 way compare.
Creative-intelligence MCP for campaign strategy, ad builds, QA, and reporting.
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.