Skip to main content
Glama
92,742 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Finding documentation and APIs for a service and using it with LLMs" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables chat-driven audio analysis and enhancement using local Claude, including denoising, EQ, compression, and loudness normalization, with an A/B viewer for synchronized comparison.
    1
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    Provides voice recognition and text extraction capabilities with support for both stdio and MCP modes, processing audio files or base64 encoded data and returning structured results with language, emotion, and speaker information.
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    This service provides fast and reliable transcriptions for audio/video files and voice memos. It allows LLMs to interact with the text content of audio/video file.
    8
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    MCP server that provides a transcribe_audio tool to convert voice messages from channels into text using OpenAI Whisper, enabling Claude Code to process audio attachments.
    1
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables spoken conversations with Claude on a Mac: Claude speaks through speakers, listens to the user's natural replies, and transcribes them locally. No audio leaves the computer, and it includes tools for voice setup and a hands-free voice mode.
    2
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Let your AI agent call your phone and talk to you — MCP servers for live, interruptible voice calls + tiered alerts, using free self-hosted pieces (pjsua2 + whisper.cpp + Linphone). No paid telephony, no extra API key.
    3
    22
    Apache 2.0
  • A
    license
    A
    quality
    A
    maintenance
    Voice interface for Claude Code: you talk, the agent listens, codes, and talks back while it works. Live speech-to-text with turn-taking, Grok/xAI voices with per-subagent personas, and a real-time HUD dashboard.
    4
    12
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    MCP server for local speech-to-text using Whisper Large V3 (MLX), enabling audio transcription with text/timestamps/SRT output and LLM-based correction, all running offline on Apple Silicon.
    2
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables generating images, video, audio, and speech from MCP clients using your own Vidofy account, with access to hundreds of models for text-to-video, image-to-video, image editing, lipsync, text-to-speech, and voice cloning.
    9
    109 npm
    1
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Enables local text-to-speech synthesis for Claude and Cursor using Supertonic 3, with support for multiple voices, expressions, and languages. No API key or cloud required.
    3
    MIT
  • A
    license
    A
    quality
    A
    maintenance
    Enables AI agents to speak using MacOS native text-to-speech, with support for blocking and non-blocking speech and a sequential queue.
    2
    2
    15
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    An MCP server that enables transcribing local audio files and Telegram voice messages using OpenAI's Whisper via local inference or cloud API. It supports multiple audio formats, automatic language detection, and optional word-level timestamps for AI-powered audio analysis.
    5
    1
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Enables MCP-capable assistants to transcribe local audio files or URLs and perform speaker diarization for Spanish and Portuguese audio, with options for speaker count hints, domain prompts, and transcript retrieval in multiple formats.
    4
    7 npm
    MIT