quick-tts-mcp
Provides text-to-speech functionality using OpenAI's TTS API, enabling conversion of text to speech with customizable voices and models.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@quick-tts-mcpconvert 'Hello, welcome to the future' to speech using voice nova"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Quick-TTS MCP Server
MCP server providing text-to-speech functionality using OpenAI's TTS API via the quick-tts package.
Installation
Using uvx (Recommended)
uvx quick-tts-mcpUsing pip
pip install quick-tts-mcpUsing Docker
docker run -e OPENAI_API_KEY=your-key quick-tts-mcpRelated MCP server: Open AI Text To Speech1 MCP Server
Configuration
Create a .env file:
OPENAI_API_KEY=sk-your-openai-api-key-hereUsage
Start the server
quick-tts-mcpAvailable tools:
generate_speech: Convert text to speechlist_voices: List available voiceslist_models: List available models
Example MCP client configuration
{
"mcpServers": {
"quick-tts": {
"command": "uvx",
"args": ["quick-tts-mcp"],
"env": {
"OPENAI_API_KEY": "sk-your-key"
}
}
}
}Tools
generate_speech
Convert text to speech using OpenAI's TTS API.
Parameters:
text(required): Text to convert to speechvoice(optional): Voice to use - alloy, echo, fable, onyx, nova, shimmer (default: alloy)model(optional): Model to use - tts-1, tts-1-hd (default: tts-1-hd)output_format(optional): Output format - mp3, wav (default: mp3)
Returns: JSON with file path, size, and metadata
list_voices
List all available TTS voices with descriptions.
Returns: JSON array of available voices
list_models
List all available TTS models with descriptions.
Returns: JSON array of available models
Development
Local installation
pip install -e .Running locally
python -m quick_tts_mcp.serverLicense
MIT
This server cannot be deployed
Maintenance
Related MCP Connectors
AI voice generation: text-to-speech and voice cloning from any MCP client.
Manage ElevenLabs voice agents and generate speech, music, sound effects, images, and video.
Text-to-speech API: neural voices, pay-per-credit in Bitcoin sats via BTCPay.
Speech, transcription, voice agents, Trace, Recap, dubbing and narration with browser OAuth.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables agents to convert text to speech using OpenAI's TTS models with voice selection, delivery instructions, and queue-based audio playback. Supports both blocking and non-blocking modes for flexible audio generation and playback control.3BSD 3-Clause
- AlicenseCqualityDmaintenanceEnables users to convert text into high-quality audio by accessing the OpenAI Text-to-Speech API. It supports customizable model selection and voice options for synthesized speech generation via the MCP protocol.1MIT
- AlicenseAqualityFmaintenanceEnables text-to-speech conversion using ElevenLabs API with voice management, streaming support, and multiple models.51MIT
- FlicenseNot gradedqualityDmaintenanceProvides text-to-speech conversion through a unified MCP interface, supporting both local Kokoro and cloud OpenAI TTS engines with streaming audio, voice selection, and customization via natural language instructions.7-