Skip to main content
Glama
91,352 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Ruby on Rails" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    A
    quality
    C
    maintenance
    Enables spoken conversations with Claude on a Mac: Claude speaks through speakers, listens to the user's natural replies, and transcribes them locally. No audio leaves the computer, and it includes tools for voice setup and a hands-free voice mode.
    2
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Image, video, chat and text-to-speech models (GPT Image, Gemini Image, Veo, Kling, Seedance, Claude, GPT, Gemini) behind one API key. generate_image returns a preview the model can see, and review_image has a vision model critique the result and propose a corrected prompt, so an assistant can generate, check and fix images on its own.
    6
    647 npm
    1
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Enables AI agents to perform local audio tasks such as speech synthesis, voice cloning, music and sound effect generation, and audio editing through MCP, with GPU models loaded on demand and released after idle.
    6
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    A cross-platform MCP server that enables Claude to speak using Microsoft Edge TTS with support for over 300 voices across 50+ languages. It requires no API keys and allows for customization of speech rate, volume, and pitch.
    3
    22 PyPI
    2
    MIT
  • A
    license
    B
    quality
    F
    maintenance
    Enables text-to-speech functionality on macOS using the say command, offering extensive control over speech parameters like voice, rate, volume, and pitch for a customizable auditory experience.
    2
    14 npm
    20
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Exposes a single transcribe tool over streamable HTTP so containerised agents can send a media file name and receive text transcribed locally by MacWhisper on the host Mac's GPU, with token-gated access and no uploads or API keys. Callers place media in a configured directory, and the blocking call returns the finished transcript.
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables AI assistants to speak with realistic cloned voices via ElevenLabs TTS on Cloudflare Workers, supporting 29 languages and inline audio playback on both desktop and mobile.
    31 npm
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    VoiceLayer MCP server enables AI coding assistants to speak and hear via local, on-device speech-to-text and text-to-speech, with no cloud dependencies.
    90 npm
    Apache 2.0
  • A
    license
    Not graded
    quality
    C
    maintenance
    Voice interface for Claude Code enabling hands-free, conversational interaction entirely on-device for Apple Silicon Macs. It provides push-to-talk transcription and automatic spoken responses via local AI models.
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    Text-to-speech MCP server using the Kokoro-82M model accelerated with MLX on Apple Silicon, enabling local Claude and Codex clients to speak text aloud and convert text to audio.
    4
    MIT
  • F
    license
    Not graded
    quality
    B
    maintenance
    Enables programmatic control of Yandex smart home from LLMs: text-to-speech on Alice speakers, batch reminders, device control, and complete scenario management.
    -
  • F
    license
    Not graded
    quality
    D
    maintenance
    Exposes a text-to-speech announcement tool for the CastleCall home PA system on Raspberry Pi. It enables users to broadcast voice messages using local Piper or cloud-based ElevenLabs providers through the Model Context Protocol.
    -