"A search engine called Brave Search" matching MCP connectors:
Matching Connector Tools:
Drive real Android & iOS devices and web browsers from natural language for mobile + web QA. 145+ tools across device control, app management, automation sessions, browser automation, and flow recording / replay. Bearer-auth — get a token at robotactions.com → Profile → API Tokens.
Prompt evals over MCP: run a prompt on your dataset, score each output 1-5 with an LLM judge.
Score any URL against a real design contract — 40 checks, A-F grade, token + motion validation.
Post-scrape data cleaner, no LLM: repairs mojibake, HTML, invisible chars. Plus a verdict.
A webhook inbox for agents: one call returns a live URL. Mock, verify, inspect and replay.
Accessibility pre-checks (WCAG/BFSG) in a real browser + statement drafts. Pay per call.
Grade MCP servers A to F with the open behavioral litmus. npm: full toolset; hosted: lookups only.
Hire a real human for real-world verification, product testing, AI output review, and errands.
Deterministic recipe verification engine — validates AI-generated recipes against master SOPs.
Runs your code against a contract; returns HELD or BROKE at the exact input. Deterministic.
MCP server for the Fail Modes taxonomy — a knowledge base of AI system failure modes
A flock of AI users tests your deployed app and reports where real people get stuck, with fixes.
Adversarial behavioural-bias engine — audits your decisions for cognitive biases via your own AI.
MCP-native AI browser testing for coding agents. Submit a URL + goal, get back action trail, bugs, screenshots, and WebM video your agent patches from directly. 43 tools, 12 AI evaluation personalities, combo tiers with auto-pause-on-bugs, throwaway email + SMS inboxes.
Rule engine with built-in simulation. 55 MCP tools for complete business rule lifecycle management.
295k+ bug-fix patterns with MCP Hub proxy, PII filtering, and code search
Pay-per-call AI evaluation MCP server. Score LLM outputs against benchmark rubrics via Workers AI.
Disposable test mailboxes on a real domain: send, receive and assert on real email.
Check if ChatGPT, Claude and Perplexity can reach, read and quote a website.
AI agent testing: replay real sessions against a rebuilt staging environment to catch regressions.