Skip to main content
Glama

Related Servers

Alternatives to quota-mcp

No user-submitted related servers found.

    Related Servers

    • F
      license
      Not graded
      quality
      B
      maintenance
      Enables AI agents and multi-agent frameworks to throttle outbound LLM API calls with multi-provider token-bucket and sliding-window limiting, plus predictive token-consumption budgeting to avoid TPM/RPM throttling. It runs deterministically over MCP with zero dependencies, returning rate-limit decisions and telemetry for any client such as Claude Desktop or Cursor.
      7
      -
    • A
      license
      Not graded
      quality
      C
      maintenance
      Routes subagent tasks across Claude and OpenAI lanes by task kind and quota state, deciding whether to spawn a single model, run a duel of both vendors, or run both and ship the merge. It records duels with attested proof-of-run session ids, grades sides blind with one judge per vendor, and tracks per-kind standings and lane pacing.
      MIT
    • A
      license
      Not graded
      quality
      B
      maintenance
      Exposes the rate-limit quota remaining and the model slugs accepted by each coding agent installed on the machine, so any MCP client can check whether work will fit before starting it, list valid models, and diagnose why a provider is not answering.
      40 npm
      MIT
    • A
      license
      Not graded
      quality
      D
      maintenance
      Provides usage data from @ccusage/pi and tools to switch between AI providers (e.g., Claude Code, OpenAI Codex, CursorAI) using load-balancing or high-availability strategies.
      14 npm
      2
      MIT
    • F
      license
      Not graded
      quality
      B
      maintenance
      Enables AI agents and MCP-compatible clients to enforce multi-provider token-bucket and sliding-window rate limits, preventing TPM/RPM throttling across LLM APIs with predictive token consumption budgeting.
      7
      -
    • A
      license
      Not graded
      quality
      A
      maintenance
      Enables agents to check remaining Anthropic and OpenAI rate-limit quotas with reset times, plus Hermes Agent token usage, so bots can throttle or stop before hitting limits.
      MIT