Skip to main content
Glama
alphaparkinc

genpark-multi-model-prompt-regression-benchmark-evaluator-skill

Related Servers

Alternatives to genpark-multi-model-prompt-regression-benchmark-evaluator-skill

No user-submitted related servers found.

    Related Servers

    • A
      license
      B
      quality
      B
      maintenance
      Enables building and running custom LLM benchmarks with multi-judge evaluation, supporting GUI, MCP client, and CLI usage for ranked, auditable results.
      3
      MIT
    • A
      license
      Not graded
      quality
      B
      maintenance
      Provides deployable, stateless MCP services for rigorous inference evaluation and benchmarking, including a profiled lm-evaluation-harness controller/worker and a bounded vLLM forward-pass benchmark adapter.
      Apache 2.0
    • A
      license
      Not graded
      quality
      C
      maintenance
      Enables benchmarking of local LLM models (performance and quality) and sharing results to a public leaderboard via MCP tools.
      15
      5
      Apache 2.0
    • A
      license
      Not graded
      quality
      A
      maintenance
      This MCP server provides a stateful, resettable, verifiable API runtime that gates every tool call, enabling agents to run long workflows against provider-shaped environments without live provider write access. It records decisions, side effects, and outcome evidence for replayable, verifiable benchmark runs.
      Apache 2.0

    Latest Blog Posts

    MCP directory API

    We provide all the information about MCP servers via our MCP API.

    curl -X GET 'https://glama.ai/api/mcp/v1/servers/alphaparkinc/genpark-multi-model-prompt-regression-benchmark-evaluator-skill'

    If you have feedback or need assistance with the MCP directory API, please join our Discord server