"Features of Grid Trading Strategies" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Risk-scan a diff, flag AI-generated-code tells, find secrets. 5 of 7 tools need no account.
Ensemble testing of web pages for accessibility, usability, and standards conformity
An open-source benchmark of how well LLMs solve nonogram puzzles, from 5x5 to 20x20.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
UI Verify is visual regression testing built for coding agents. Connect the MCP server and your agent (Claude Code, Cursor, Codex) pulls a pull request's UI changes into the conversation, views each visual diff, reads the AI judge's verdict of regression vs intended change, and accepts the intended baselines - all over MCP.
Check whether AI assistants can reach, read and cite a public website. No account for 4 of 5 tools.
A digital audit for small-business websites: a score out of 100 and the main problems.
AI users run real tasks on your live site and show where they get stuck, with a replay of every step
Deterministic check of logged service hours: bad dates, impossible totals, duplicates. No AI.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
Compare two versions of a JSON row list: what was added, removed or changed, field by field.
Checks the structural integrity of translated resource dictionaries against a source dictionary y...
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
MCP server for static security analysis of Android source code
Reusable checklists and dated runs of them: pass, fail, not applicable, and a sign-off.
Evidence-bound second-opinion audit of an agent conclusion against caller-supplied evidence.
- CurrentsOAuth unavailabledev.currents
Self-healing CI tests and proof of work for AI agents
Document reasoning checks: a helicopter view of claims and assumptions before handoff.
Probability of Backtest Overfitting (CSCV), Deflated Sharpe Ratio, and purged CV splits.