Skip to main content
Glama

Related Servers

Alternatives to mcp-turboquant

No user-submitted related servers found.

    Related Servers

    • A
      license
      Not graded
      quality
      C
      maintenance
      A local, zero-cloud MCP server for token and text compression. It provides tools to compress, auto-compress, measure, and decompress text using offline rules, lossless gzip packing, or a local Ollama semantic model.
      1
      MIT
    • A
      license
      A
      quality
      B
      maintenance
      MCP server for QuelLLM: recommends the best open-source LLM to run locally for your hardware (GPU/RAM), with model comparison and a cost calculator.
      6
      MIT
    • A
      license
      Not graded
      quality
      C
      maintenance
      MCP server for compressing AI embeddings by 5-7x using TurboQuant (PolarQuant + QJL), with tools to compress, decompress, estimate savings, and embed+compress vectors.
      MIT
    • A
      license
      A
      quality
      A
      maintenance
      Unified MCP server for managing local model runtimes (Ollama, LM Studio, etc.), enabling provider-agnostic discovery, lifecycle management, hardware-fit checks, and delegated inference.
      16
      18 npm
      Creative Commons Attribution Non Commercial No Derivatives 4.0 International
    • F
      license
      B
      quality
      D
      maintenance
      Local MCP server for token optimization, providing tools to compress code/JSON, optimize prompts, and manage placeholder-based content redaction and hydration to reduce LLM token usage.
      5
      -

    TDQS

    A4.4/5.0

    Scored across 6 tools

    Disambiguation5/5

    Each tool targets a distinct action in the quantization workflow: system check, model info, recommendation, quantization, evaluation, and upload. No overlap or ambiguity.

    Naming Consistency4/5

    All tools use single imperative verbs (check, evaluate, info, push, quantize, recommend). 'Info' is a noun rather than verb, but the pattern is otherwise consistent and clear.

    Tool Count5/5

    6 tools cover the entire quantization lifecycle without bloat or missing essentials. Each tool earns its place for a focused server purpose.

    Completeness4/5

    Covers system check, model info, recommendation, quantization, evaluation, and upload. Minor gap: no tool to list previously quantized models locally, but core workflow is complete.

    Maintenance

    ActivityInactive
    ResponsivenessNo issues