feature-separate-batch-eval-mcp
Related Servers
Alternatives to feature-separate-batch-eval-mcp
No user-submitted related servers found.
Related Servers
- FlicenseAqualityBmaintenanceEnables evaluation of patent feature-separation results by listing records, checking alignment with claims and original text, and running deterministic quality gates plus a Judge audit.3-
- AlicenseNot gradedqualityAmaintenanceEnables conversational patent data analysis through a set of tools for trend analysis, clustering, word cloud generation, and more. Allows natural language queries to explore patent datasets with structured results and visualizations.1AGPL 3.0
- AlicenseAqualityDmaintenanceEnables analysis of clinical trial protocols using MCP tools for document listing, entity extraction, adverse event clustering, and summarization.41MIT

Patronus MCP Serverofficial
AlicenseNot gradedqualityDmaintenanceEnables running LLM evaluations, experiments, and custom evaluators through a standardized MCP interface.16Apache 2.0- AlicenseNot gradedqualityBmaintenanceEnables evidence-first game operations incident investigation through read-only MCP tools, including metric queries, cohort comparisons, anomaly detection, and reproducible incident report drafting with citations.76MIT
- AlicenseAqualityBmaintenanceEnables analyzing coding-agent session transcripts to measure tool efficiency, detect loops, track test outcomes, and report costs, with MCP tools for listing sessions, analyzing sessions, finding loops, and generating cost reports.4MIT
TDQS
Scored across 5 tools
Each tool targets a distinct action: status (list cohorts), tag (assign records), check (validate one cohort), run (execute network eval), insights (list persisted results). Status and check both touch 'readiness' and status/insights both list things, which introduces mild overlap, but descriptions keep them separable.
Every tool shares the consistent 'feature_separate_batch_eval_' prefix followed by a short suffix, forming a predictable pattern. Minor deviation in that tag/check/run are verbs while status/insights are nouns, but overall highly consistent.
Five tools is a reasonable, well-scoped set for a batch-evaluation workflow (readiness, tagging, validation, execution, insights). Slightly lean but each tool has a clear role without redundancy.
The core lifecycle (list readiness, assign cohorts, validate, run, retrieve insights) is covered, but there is no way to modify/remove cohort assignments or clear/export persisted insights, leaving some dead ends.