ElevenLabs MCP Server
Provides comprehensive access to ElevenLabs API features including text-to-speech conversion, voice management, speech-to-speech transformation, sound effect generation, audio isolation for noise removal, and user account management with history tracking.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@ElevenLabs MCP Serverconvert 'Welcome to our meeting' to speech using a friendly voice"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
ElevenLabs MCP Server
A comprehensive Model Context Protocol (MCP) server for the ElevenLabs API, providing access to all major ElevenLabs features including text-to-speech, voice generation, audio isolation, and more.
Features
This MCP server provides tools for:
Text to Speech
text-to-speech: Convert text to speech with customizable voice settings
text-to-speech-streaming: Stream text-to-speech audio in real-time
Voice Management
get-voices: List all available voices with search and filtering
get-voice: Get detailed information about a specific voice
get-models: List all available AI models
Audio Transformation
speech-to-speech: Transform audio from one voice to another (voice changer)
sound-generation: Create sound effects from text descriptions
audio-isolation: Remove background noise from audio
History & User Management
get-history: Get history of all generated audio
get-history-item: Get a specific history item by ID
get-history-item-audio: Download audio from a history item
delete-history-item: Delete a history item
get-user: Get current user information
get-subscription: Get user subscription details
Related MCP server: ElevenLabs MCP Server
Installation
npm installConfiguration
The server requires an ElevenLabs API key. You can configure this when connecting the server to your MCP client.
Configuration Schema
{
"apiKey": "your-elevenlabs-api-key-here"
}Usage with Smithery
Development
npm run devThis will start the server in development mode with hot reloading.
Build
npm run buildDeploy to Smithery
Push your code to GitHub
Go to Smithery
Click "Deploy" and connect your GitHub repository
Usage with Claude Desktop or Cursor
Add this to your MCP settings:
{
"mcpServers": {
"elevenlabs": {
"url": "your-smithery-deployment-url",
"config": {
"apiKey": "your-elevenlabs-api-key"
}
}
}
}API Key
Get your ElevenLabs API key from ElevenLabs Settings.
Example Usage
Generate Speech
Use the text-to-speech tool to convert "Hello, world!" to speech using voice ID "21m00Tcm4TlvDq8ikWAM"List Available Voices
Use the get-voices tool to see all available voicesCreate Sound Effect
Use the sound-generation tool to create a "dog barking" sound effectRemove Background Noise
Use the audio-isolation tool to remove background noise from an audio file (provide base64 encoded audio)API Reference
All tools follow the ElevenLabs API documentation.
Audio Format
Audio files are returned as base64-encoded strings. Supported formats include:
MP3 (various bitrates)
PCM (various sample rates)
μ-law format (for Twilio)
Development
The server is built using:
@modelcontextprotocol/sdk - MCP SDK
@smithery/sdk - Smithery SDK
zod - Schema validation
License
MIT
Resources
This server cannot be deployed
Maintenance
Related MCP Connectors
Manage ElevenLabs voice agents and generate speech, music, sound effects, images, and video.
ElevenLabs in natural language: generate speech in any language, create and manage voices, compose m
Audio for your agent: transcribe, speak, translate, summarise, plus sound effects and music.
Audio AI tools: text-to-speech, voice cloning, music generation, stem separation, transcription.
Related MCP Servers
- AlicenseAqualityNot gradedmaintenanceEnables interaction with ElevenLabs Text-to-Speech and audio processing APIs. Supports speech generation, voice cloning, audio transcription, and sound effect creation through natural language.24MIT
- AlicenseAqualityFmaintenanceEnables text-to-speech conversion using ElevenLabs API with voice management, streaming support, and multiple models.51MIT
- FlicenseAqualityDmaintenanceEnables text-to-speech audio generation using ElevenLabs voices directly from Claude conversations, supporting single and batch conversion, voice listing, and voice ID lookup.5-
- AlicenseBqualityBmaintenanceEnables interaction with ElevenLabs AI models (audio isolation, speech-to-text, text-to-dialogue, sound effects, text-to-speech) through RunAPI, supporting task creation, status polling, and pricing checks.889 npmApache 2.0