eval_discover
Return a machine-readable catalog of evaluators, trap families, suites, calibration thresholds, and versions to plan an evaluation strategy without guessing tool names.
Instructions
Return the full machine-readable capability catalog.
Useful as a first call at session start — an agent can plan its evaluation strategy against the actual available evaluators rather than guessing or hallucinating tool names.
Returns: A dict with six top-level keys:
- ``evaluators``: every available multivon-eval evaluator,
with its category, import path, evaluator ID, and summary.
- ``traps``: every pdfhell trap family, the failure mode each
elicits, and the expected_failure_mode metadata.
- ``suites``: every named pdfhell suite, the (trap_family,
seed_count) breakdown, and the suite_hash for the canonical
version.
- ``calibration``: shipped per-(evaluator, judge) threshold rows.
- ``version``: installed multivon-mcp, multivon-eval, and pdfhell
versions.
- ``server``: the stable server identifier.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||