Skip to main content
Glama
97,619 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Edge Impulse" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    Not graded
    quality
    B
    maintenance
    Provides an MCP-compatible profiler for measuring edge and on-device LLM inference latency, including TTFT, TPOT, tokens-per-second throughput, and P50/P90/P99 jitter distributions. It enables agents and developers to run these telemetry benchmarks from MCP clients such as Claude Desktop, Cursor, and Windsurf.
    7
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    From idea to edge. Everyone has unique opinions and views of the world. Varrd makes it possible for everyone regardless of statistical, coding, or market knowledge to be able to find their unique edge. The issue with LLMs testing for edges in the market is redundant idea loops, overfit with confidence, and waste days exploring nonsense. Varrd is the infrastructure and guardrails to prevent those f
    9
    12,729 PyPI
    24
    MIT
  • A
    license
    B
    quality
    B
    maintenance
    An ultra-rational A2A protocol for zero-token edge pre-filtering and FEP-driven deadlock prevention. Uses Cloudflare Vectorize (384d cosine similarity) with a 24h deposit model, restricting bargaining to a 4-rally limit before forcing HTTP 402 dimension jumps.
    2
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables zero-dependency INT8 symmetric quantization of LLM weights and activations on a per-tensor or per-channel basis, reporting SNR and MSE reconstruction telemetry along with TTFT/TPOT latency and jitter metrics. Also exposes edge inference primitives such as PagedAttention block allocation, radix prefix caching, and speculative decoding verification through the Model Context Protocol.
    7
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables MCP clients to manage LLM inference memory through virtual paged-attention KV-cache block mapping, non-contiguous physical page allocation, zero-copy fragmentation tracking, radix-trie prefix caching, INT8 quantized compute, and speculative decoding verification. It also exposes prefill and decode latency telemetry so edge deployments can be benchmarked and tuned without external dependencies.
    7
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    Provides MCP tools for edge-based pharmaceutical shipment disposition decisions, returning release, review, or quarantine outcomes while recording every decision in a tamper-evident, hash-chained audit trail.
    Academic Free v1.1
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables analysts to build and evaluate simultaneous VARMA models for systems of time series via exact maximum likelihood on the ATSW ladder, using univariate models as the yardstick. It guides users through loading univariate fits, identifying residual cross-correlations, estimating joint models, forecasting, impulse responses, variance decomposition, and recording decisions.
    GPL 2.0
  • A
    license
    Not graded
    quality
    A
    maintenance
    Enables AI agents to delegate simple text subtasks such as summarising, translating, classifying, extracting, and reformatting to free web AIs through the user's own Edge/Chrome browser, returning only plain-text answers to save the main model's context and tokens.
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables verification of draft-target speculative decoding with rejection sampling, acceptance-rate telemetry, and dynamic speedup estimation, alongside edge inference primitives such as INT8 quantization, paged KV allocation, and radix prefix caching.
    8
    MIT