Skip to main content
Glama
95,407 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"MCP servers with streamable-http support" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables chat-driven audio analysis and enhancement using local Claude, including denoising, EQ, compression, and loudness normalization, with an A/B viewer for synchronized comparison.
    1
    MIT
  • A
    license
    Not graded
    quality
    C
    maintenance
    Self-hosted WhatsApp management over Model Context Protocol, exposing a streamable HTTP MCP endpoint with 30 tools for session/QR pairing, messaging, media storage, and optional on-CPU voice note transcription via whisper.cpp.
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Enables continuous voice conversation with AI coding assistants by locally transcribing speech with Whisper and delivering utterances as text prompts.
    4
    1
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    MCP server that provides a transcribe_audio tool to convert voice messages from channels into text using OpenAI Whisper, enabling Claude Code to process audio attachments.
    1
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Let your AI agent call your phone and talk to you — MCP servers for live, interruptible voice calls + tiered alerts, using free self-hosted pieces (pjsua2 + whisper.cpp + Linphone). No paid telephony, no extra API key.
    3
    23
    Apache 2.0
  • A
    license
    A
    quality
    C
    maintenance
    MCP server for local speech-to-text using Whisper Large V3 (MLX), enabling audio transcription with text/timestamps/SRT output and LLM-based correction, all running offline on Apple Silicon.
    2
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables generating images, video, audio, and speech from MCP clients using your own Vidofy account, with access to hundreds of models for text-to-video, image-to-video, image editing, lipsync, text-to-speech, and voice cloning.
    9
    82 npm
    2
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Provides on-device Chinese speech recognition with traditional Chinese (Taiwan) output via MCP server, HTTP API, and CLI, ensuring privacy by processing audio locally without uploading to cloud services.
    2
    1
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Provides AI tools with transcripts of any YouTube video, including videos without captions via AI speech-to-text, with language selection and length capping.
    1
    126 npm
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables AI agents to access a local Anki collection through AnkiConnect: list decks, search and retrieve notes/cards, identify weak cards, and batch-add validated flashcards with optional pronunciation audio.
    5
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Enables local text-to-speech synthesis for Claude and Cursor using Supertonic 3, with support for multiple voices, expressions, and languages. No API key or cloud required.
    3
    MIT