Skip to main content
Glama

Related Servers

Alternatives to judge-audit-mcp

No user-submitted related servers found.

    Related Servers

    • A
      license
      A
      quality
      B
      maintenance
      An MCP server that lets any AI agent evaluate RAG outputs -- faithfulness scoring, hallucination detection, and retrieval quality metrics -- with zero API keys, using MCP sampling.
      6
      MIT
    • F
      license
      Not graded
      quality
      C
      maintenance
      MCP server that scans multi-language codebases to detect stubs, missing imports, and structural incompleteness, assigning drift scores to quantify LLM context erosion.
      -
    • A
      license
      B
      quality
      C
      maintenance
      An MCP server implementing the Flourishing-Justice-Autonomy (FJA) alignment framework, enabling FJA evaluation and fine-tuning of LLMs through examples like cultural diet, medical autonomy, and hiring fairness.
      2
      MIT

    TDQS

    A4.2/5.0

    Scored across 6 tools

    Disambiguation5/5

    Each tool serves a clearly distinct purpose: comprehensive audit, specific bias testing, calibration against humans, drift detection, data generation, and metric explanation. No two tools overlap in functionality.

    Naming Consistency5/5

    All tool names follow a consistent verb_noun pattern in snake_case (e.g., audit_judge, bias_probe, explain_metric). The convention is uniform and predictable.

    Tool Count5/5

    6 tools cover the essential aspects of LLM judge auditing without redundancy or bloat. Each tool is justified and serves a specific role in the workflow.

    Completeness4/5

    The set covers comprehensive auditing, bias probes, calibration, drift detection, data generation, and explanation. Minor gap: no tool for running a judge on raw inputs, but that is likely external to this server's scope.

    Maintenance

    ActivityStale
    ResponsivenessNo issues