Text-to-Speech
Tools for converting text-to-speech and vice-versa.
MCP ServersBrowse all →
AlicenseAqualityAmaintenanceVoice interface for Claude Code: you talk, the agent listens, codes, and talks back while it works. Live speech-to-text with turn-taking, Grok/xAI voices with per-subagent personas, and a real-time HUD dashboard.649MIT- AlicenseAqualityAmaintenanceEnables AI agents to speak using MacOS native text-to-speech, with support for blocking and non-blocking speech and a sequential queue.2215MIT

@lumiastream/mcpofficial
AlicenseAqualityDmaintenanceAn MCP server for Lumia Stream that lets AI assistants trigger commands, set light colors, fire alerts, run text-to-speech, and read/write variables via the local REST API.4120 npmMIT
@vocea.app/mcp-serverofficial
AlicenseAqualityCmaintenanceEnables AI agents to generate speech, transcribe audio, and manage voices via the Vocea API.6MIT- AlicenseAqualityAmaintenanceGive your AI agent a voice with x402 pay-per-call speech synthesis, offering 20 voices, 10 personas, 31 languages, and granular controls.4666 npmMIT
- AlicenseAqualityAmaintenanceText to speech for MCP clients. Reads numbers, dates and order IDs correctly. 23 languages, six voices, every render watermarked. Free key with 100,000 characters, no card.4MIT
- Apache 2.0

@paxalabs/mcpofficial
AlicenseAqualityBmaintenanceEnables agents to speak Thai and English audio through local speakers, manage a playback queue, save speech to files, translate text into Thai, and OCR PDFs and images.101,249 npm10MIT
MiniMax MCP Serverofficial
AlicenseAqualityCmaintenanceEnables MCP clients like Claude Desktop and Cursor to interact with MiniMax APIs for generating speech, cloning voices, creating videos, and generating images.61,577MIT
MiniMax MCP JSofficial
AlicenseAqualityDmaintenanceJavaScript implementation of MiniMax MCP that enables interaction with MiniMax AI services for image generation, video generation, text-to-speech, and voice cloning through MCP-compatible clients.10252 npm128MIT- AlicenseAqualityDmaintenanceAI-powered multi-voice audiobook creation platform. Provides tools for pricing, language support, use cases, onboarding, FAQ, alternatives comparison, and cost estimation. npx echo3s-mcp129 npmMIT

supertone-mcpofficial
AlicenseAqualityFmaintenanceMCP server for the Supertone TTS API. Generate natural speech, browse and preview the voice catalog, predict synthesis cost, and create cloned voices — directly from Claude Desktop, Cursor, or any MCP-compatible client. Supports Korean, English, Japanese, and 20+ other languages, with speed, pitch, and emotion-style control.144MIT- AlicenseBqualityDmaintenanceA Model Context Protocol server that provides text-to-speech capabilities using the Kokoro TTS model, offering multiple voice options and customizable speech parameters.424 npm1MIT
- AlicenseAqualityNot gradedmaintenanceEnables interaction with MiniMax AI APIs for text-to-speech, voice cloning, video generation, image generation, and music creation through MCP clients like Claude Desktop and Cursor.9MIT
- AlicenseAqualityBmaintenanceEnables text-to-speech and speech-to-text through MCP tools, using Groq's free hosted endpoints when configured and falling back to fully local keyless models otherwise. Also provides voice listing and provider health checks.4MIT
- AlicenseBqualityDmaintenanceEnables natural language-driven speech synthesis using Fish Audio's Text-to-Speech API, supporting multiple voices, streaming, and flexible configuration.217 npmMIT
- AlicenseAqualityBmaintenanceEnables MCP clients to synthesize text into speech locally and play it through the machine's audio, with selectable voices and playback speed.2Apache 2.0
- AlicenseAqualityAmaintenanceText-to-speech MCP server that enables AI assistants to read text aloud on the user's computer using Windows SAPI, with no API key or cloud service required.11MIT
- AlicenseAqualityDmaintenanceProvides high-quality text-to-speech synthesis with 10 natural voices, emotion control, and dynamic pacing for professional applications requiring expressive speech output.52MIT
- AlicenseAqualityCmaintenanceMCP server for text-to-speech using macOS say command, enabling speech synthesis, audio file generation, and voice management.56 npm1MIT
- AlicenseAqualityNot gradedmaintenanceEnables interaction with ElevenLabs Text-to-Speech and audio processing APIs. Supports speech generation, voice cloning, audio transcription, and sound effect creation through natural language.24MIT
- AlicenseAqualityDmaintenanceAn MCP server that enables LLMs to generate spoken audio from text using OpenAI's Text-to-Speech API, supporting various voices, models, and audio formats.16 npm1MIT
- AlicenseAqualityFmaintenanceEnables AI assistants to access Venice AI's capabilities including chat with open-source models, image generation, text-to-speech, embeddings, and API key management.1321 npm5MIT
- AlicenseAqualityCmaintenanceA simple MCP server that can send notifications on mac devices.59 npm26MIT
- AlicenseBqualityFmaintenanceA feature-rich MCP server for Discord, giving AI companions full presence in Discord communities — reading, responding, reacting, searching, and speaking.26411 npm1MIT
- AlicenseAqualityCmaintenanceEnables text-to-speech synthesis through a streaming, GPU-accelerated gateway, exposing a single tool that returns playable WAV audio with configurable voice and speed.1MIT
- AlicenseAqualityFmaintenanceIntegrates with ElevenLabs text-to-speech API.6118MIT
- AlicenseAqualityCmaintenanceExposes the canonical WordCast knowledge surface — voice and TTS workflows, blog topics, FAQ, official links — to MCP-compatible AI clients. Read-only, no API keys required.2MIT
- AlicenseAqualityCmaintenanceEnables any MCP-compatible client to perform local voice cloning and text-to-speech using OmniVoice, including voice profile management, voice design, and an optional ASMR post-processing DSP pipeline.6MIT
- AlicenseBqualityDmaintenanceProvides AI-powered audio generation and processing through the MusicGPT API, enabling music creation, voice conversion, audio manipulation, stem extraction, and audio analysis capabilities.249 npm1MIT
MCP ConnectorsBrowse all →
APICK Korean data, OCR, search, conversion, image and video generation, and asynchronous TTS
APKBuild.org services: convert websites to Android APK/.aab and AI voiceover MP3.
File conversion: PDF, DOCX, STT, TTS, watermarking
Bambara AI over MCP: text-to-speech, transcription and translation (Bamanankan + more).
Your AI rings your iPhone, speaks its question, and gets your spoken answer back as text.
Free IELTS prep: band-scored student essays and interactive Listening/Reading drills graded in-chat.
Kurdish (Sorani & Kurmanji) text-to-speech & speech-to-text — 664 AI voices. API key required.
ElevenLabs in natural language: generate speech in any language, create and manage voices, compose m
Free receptionist tools: phone scripts, IVR menus (EN+ES), ElevenLabs prompts, missed-call math
Curated audio news, daily briefings, the Declassified library + market-linked signals.
Text to speech for your AI. Your AI can send text to Doc Player to read it aloud. You will see a reader window with the text and you can control the playback sentence by sentence. Find an example here: https://documentplayer.com/connect-ai/
Generate highly realistic Text to Speech voiceovers.
Audit localized tutorials and safely maintain project metadata and pronunciation rules.
Hosted pay-per-use TTS: 54 neural voices, 9 languages incl. Brazilian Portuguese. $10 free credits.
Manage Speko voice-AI agents, sessions, calls, phone numbers, knowledge bases, evals, and docs.
- ElevenLabsOAuth unavailableio.elevenlabs
Manage ElevenLabs voice agents and generate speech, music, sound effects, images, and video.
Hosted MCP for speech cleanup. Remove noise, filler words, and dead air. Also cut, join, transcribe, and TTS. Sign in with a CleanAudio account. No API key. Endpoint: https://mcp.cleanaudio.app/mcp This is cleanaudio.app, not cleanaudio.io.
Polish SMS & voice gateway: send SMS/TTS, contacts, blacklist, tracked links, replies, reports.
Pronunciation scoring, speech-to-text, and text-to-speech for language learning
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning)