"Help with coding on Cursor IDE" matching MCP connectors:
Matching Connector Tools:
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Evidence-gated task verification for AI agents. Decompose goals into acceptance criteria, attach proof (screenshot, curl, file), independent LLM judge accepts or rejects. 24 tools. Hosted remote MCP (streamable-http, OAuth 2.1 + DCR).
Free platform to test MCP clients without installing anything. Create mock tools with dynamic templates, configurable delays, conditions (if/then), and response sequences. Supports JSON-RPC 2.0 over Streamable HTTP. Built-in text_echo and json_echo tools. Rate-limited tiers: anonymous (5 calls/min, 1 mock tool), registered (10 calls/min, 4 mock tools), premium (60 calls/min, unlimited). Zero setup — no install, no registration required. More info: https://www.testmcp.dev
Give AI coding agents access to your Vynix visual feedback, bug reports, and AI diagnosis.
Estimated game fps for any GPU or Apple Silicon chip, with the limiter and tweaks.
Voice-powered bug reporting with 13 MCP tools. Record bugs by talking; let AI find and fix them.
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
Pay-per-call MCP server. Vetted human experts review AI-generated content (text, images, video, audio, social posts), audit reasoning chains, run prepublish safety checks, and gate high-stakes actions for human approval. 7 paid tools at $1.00 each plus 4 free tools (list_offerings, list_expert_profiles, get_result, verify_certificate). Payment via x402, USDC on Base mainnet. Approved outputs receive an on-chain Taste content certificate downstream agents verify before consuming.
Benchmark-first release surface with a read-only MCP endpoint and operator CLI.
MCP-native AI browser testing for coding agents. Submit a URL + goal, get back action trail, bugs, screenshots, and WebM video your agent patches from directly. 43 tools, 12 AI evaluation personalities, combo tiers with auto-pause-on-bugs, throwaway email + SMS inboxes.
The world's first named AI prompt quality score. Score, optimize, and compare LLM prompts before they hit any model. Free tier available. Built on PEEM, RAGAS, G-Eval, and MT-Bench frameworks. x402-native on Base.
Evaluates UI designs for WCAG accessibility issues automated scanners miss. Paid via x402 on Base.
Test-inbox API for email and SMS: create inboxes, long-poll messages, extract OTPs and links. 34 tools; bearer auth with an mfx_ API key (free tier).
Rule engine with built-in simulation. 55 MCP tools for complete business rule lifecycle management.
295k+ bug-fix patterns with MCP Hub proxy, PII filtering, and code search
Evaluate, benchmark, and simulate AI agents on the VerifyAX agent-evaluation platform.
Control real Android and iOS devices with LLM agents — tap, swipe, type, automate flows.
AI dev tools + image generation, paid per-use with USDC on Base (x402).