quick-tts-mcp
Provides text-to-speech functionality using OpenAI's TTS API, enabling conversion of text to speech with customizable voices and models.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@quick-tts-mcpconvert 'Hello, welcome to the future' to speech using voice nova"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Quick-TTS MCP Server
MCP server providing text-to-speech functionality using OpenAI's TTS API via the quick-tts package.
Installation
Using uvx (Recommended)
uvx quick-tts-mcpUsing pip
pip install quick-tts-mcpUsing Docker
docker run -e OPENAI_API_KEY=your-key quick-tts-mcpRelated MCP server: Open AI Text To Speech1 MCP Server
Configuration
Create a .env file:
OPENAI_API_KEY=sk-your-openai-api-key-hereUsage
Start the server
quick-tts-mcpAvailable tools:
generate_speech: Convert text to speechlist_voices: List available voiceslist_models: List available models
Example MCP client configuration
{
"mcpServers": {
"quick-tts": {
"command": "uvx",
"args": ["quick-tts-mcp"],
"env": {
"OPENAI_API_KEY": "sk-your-key"
}
}
}
}Tools
generate_speech
Convert text to speech using OpenAI's TTS API.
Parameters:
text(required): Text to convert to speechvoice(optional): Voice to use - alloy, echo, fable, onyx, nova, shimmer (default: alloy)model(optional): Model to use - tts-1, tts-1-hd (default: tts-1-hd)output_format(optional): Output format - mp3, wav (default: mp3)
Returns: JSON with file path, size, and metadata
list_voices
List all available TTS voices with descriptions.
Returns: JSON array of available voices
list_models
List all available TTS models with descriptions.
Returns: JSON array of available models
Development
Local installation
pip install -e .Running locally
python -m quick_tts_mcp.serverLicense
MIT
This server cannot be deployed
Maintenance
Related MCP Connectors
AI voice generation: text-to-speech and voice cloning from any MCP client.
Manage ElevenLabs voice agents and generate speech, music, sound effects, images, and video.
Speech, transcription, voice agents, Trace, Recap, dubbing and narration with browser OAuth.
AI content generation with 50+ models: image, video, TTS, voice cloning, and more.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables agents to convert text to speech using OpenAI's TTS models with voice selection, delivery instructions, and queue-based audio playback. Supports both blocking and non-blocking modes for flexible audio generation and playback control.3BSD 3-Clause
- AlicenseCqualityDmaintenanceEnables users to convert text into high-quality audio by accessing the OpenAI Text-to-Speech API. It supports customizable model selection and voice options for synthesized speech generation via the MCP protocol.1MIT
- AlicenseAqualityFmaintenanceEnables text-to-speech conversion using ElevenLabs API with voice management, streaming support, and multiple models.51MIT
- FlicenseNot gradedqualityDmaintenanceProvides text-to-speech conversion through a unified MCP interface, supporting both local Kokoro and cloud OpenAI TTS engines with streaming audio, voice selection, and customization via natural language instructions.7-