"Using memory in programming with the Cursor system" matching MCP connectors:
Matching Connector Tools:
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Free platform to test MCP clients without installing anything. Create mock tools with dynamic templates, configurable delays, conditions (if/then), and response sequences. Supports JSON-RPC 2.0 over Streamable HTTP. Built-in text_echo and json_echo tools. Rate-limited tiers: anonymous (5 calls/min, 1 mock tool), registered (10 calls/min, 4 mock tools), premium (60 calls/min, unlimited). Zero setup — no install, no registration required. More info: https://www.testmcp.dev
Check that your AI is being logical. Free tool that mathematically catches contradictions in agent reasoning. No account needed. Also offers paid guardrails that converts natural language to formal verification proofs, that anyone can check succinctly.
Design-system contract verification, scoring, and review tools for AI agents.
Agentic code review, no signup to try: reality gates + frontier-model review, with veto.
Verify scraped data against the live source page. Signed verdicts, $0.01 via x402.
Real-browser WCAG audit that also finds keyboard-inoperable controls axe-core misses, with fixes.
Voice-powered bug reporting with 13 MCP tools. Record bugs by talking; let AI find and fix them.
Give AI coding agents access to your Vynix visual feedback, bug reports, and AI diagnosis.
Estimated game fps for any GPU or Apple Silicon chip, with the limiter and tweaks.
Check if your MCP server is ready to publish on the MCP Registry, Smithery, or npm.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
Test the voice agents you run: scored transcripts, pass/fail verdicts, latency and WER metrics.
Read-only MCP server for the OPERANT AI operating-agent calibration benchmark.
Machine-readable taxonomy of 100+ AI system failure modes spanning factuality, alignment, planning, code generation, and instruction following.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
PQS scores any prompt before the model runs. 8 dimensions. 5 frameworks. Pre-flight, not post-hoc.
Benchmark-first release surface with a read-only MCP endpoint and operator CLI.