Skip to main content
Glama
Mhdd-24

@mhdd_24/ai-benchmark-mcp

by Mhdd-24

Related Servers

Alternatives to @mhdd_24/ai-benchmark-mcp

No user-submitted related servers found.

    Related Servers

    • A
      license
      Not graded
      quality
      C
      maintenance
      Enables benchmarking of local LLM models (performance and quality) and sharing results to a public leaderboard via MCP tools.
      6 npm
      8
      Apache 2.0
    • A
      license
      A
      quality
      C
      maintenance
      Enables users to benchmark AI models across English, Urdu, and Roman Urdu by running question sets, grading answers, measuring confidently wrong and hedged responses, and producing reports, comparisons, and CSV exports through an MCP client or CLI.
      8
      MIT

    TDQS

    C2.8/5.0

    Scored across 4 tools

    Disambiguation4/5

    The four tools are mostly distinct: status shows configuration/health, list enumerates entities, inspect examines a single artifact, and compare contrasts two items. There is slight potential confusion between 'inspect' and 'compare' since both involve examining artifacts, but their purposes are clear enough.

    Naming Consistency4/5

    All tools share the 'aibench_' prefix and use simple verb-based names (status, list, inspect, compare). This is consistent, though the prefix is a bit verbose and 'aibench_status' is more of a noun than a verb_noun pattern.

    Tool Count4/5

    Four tools is a reasonable, focused set for a benchmark server. It covers the core operations without being bloated, though it is on the smaller side.

    Completeness3/5

    The set covers status, listing, inspection, and comparison, which are the main read-only operations for benchmarks. However, there are no tools to run, create, or delete benchmarks, so the surface is incomplete for a full lifecycle.

    Maintenance

    ActivityMaintained
    ResponsivenessNo issues