Voice MCP
The Voice MCP server provides pay-per-call speech synthesis via the https://voice.forgemesh.io API, supporting 20 voices, 10 personas, and 31 languages. Payments are made in USDC on the Base network; no API keys or subscriptions are needed.
list_voice_catalog(Free): Discover all available voices, personas, language codes, pricing tiers, character limits, and speed/quality controls before making any paid call.generate_standard_voice($0.001–$0.003): Convert text (up to 2,000 characters) to WAV audio using 10 standard voices (M1–M5,F1–F5) across 31 languages. Ideal for simple narration, alerts, and status updates.generate_controlled_voice($0.003–$0.006): Produce WAV speech with granular control over speed (0.7x–2.0x) and quality steps (1–100) for precise pacing or polished audio fidelity.generate_persona_voice($0.005–$0.01): Synthesize expressive WAV audio using 10 distinct persona voices (e.g., Storyteller, Narrator, Announcer, Assistant, Urgent, Velvet, Echo), with speed and quality controls. Best for branded agents, characters, and premium experiences.generate_openai_compatible_voice($0.001–$0.003): Submit speech requests using an OpenAI/v1/audio/speech-shaped payload (input,voice,model,response_format), supporting WAV, FLAC, and OGG output. Ideal for apps already built around the OpenAI audio API format.generate_batch_voices($0.002–$0.005): Process up to 20 texts in a single call with per-item voice and language settings, perfect for notification queues, scripted sequences, and multi-step agent workflows.
All audio is returned as audio_base64-encoded output. Input validation (voice names, language codes, character limits, speed/quality ranges) happens locally before any paid call is made.
Provides tools for generating speech using an OpenAI-compatible API, enabling text-to-speech with multiple voices, personas, languages, speed, quality controls, and batch generation.
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@Voice MCPspeak 'Hello, how are you?' in F1 voice"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Voice MCP
Give Your Agent A Voice: x402 pay-per-call speech with 20 voices, 10 personas, 31 languages, granular speed and quality controls, OpenAI-shaped requests, and batch audio.
This MCP wraps https://voice.forgemesh.io, an x402 Voice API with standard voices, persona voices, OpenAI-shaped speech requests, 31 languages, speed controls, quality controls, and batch generation. Payments are made per call in USDC on Base.
Voice Coverage
10 standard voices:
M1-M5,F1-F510 persona voices:
Storyteller,Narrator,Announcer,Assistant,Urgent,Sage,Spark,Anchor,Velvet,Echo31 languages:
en,ko,ja,ar,bg,cs,da,de,el,es,et,fi,fr,hi,hr,hu,id,it,lt,lv,nl,pl,pt,ro,ru,sk,sl,sv,tr,uk,viGranular control: speed
0.7x-2.0x, quality steps1-100, persona selection, OpenAI-shaped audio format requests, and batch generation for up to 20 textsVoice samples are generated on demand by the paid speech tools and returned as
audio_base64WAV output
Related MCP server: Spix
Voice Samples
Tools
Tool | Price | Purpose |
| Free | Voices, personas, languages, pricing, buckets, and controls |
| $0.001 / $0.003 | Low-cost speech with 10 standard voices |
| $0.003 / $0.006 | Speech with granular speed and quality controls |
| $0.005 / $0.01 | Storyteller, Velvet, Narrator, Announcer, Assistant, Urgent, and more |
| $0.001 / $0.003 | OpenAI-shaped |
| $0.002 / $0.005 | Up to 20 texts per call |
Short prices apply to 1-500 characters. Long prices apply to 501-2000 characters.
The MCP validates voice names, language codes, audio formats, speed/quality ranges, batch item count, and character limits locally before making a paid x402 call.
Install
npm install -g @forgemeshlabs/voice-mcpDocker
Build:
docker build -t voice-mcp .Run over stdio:
docker run --rm -i \
-e WALLET_PRIVATE_KEY=0x... \
voice-mcpCMD arguments:
["node", "index.js"]MCP Config
{
"mcpServers": {
"voice": {
"command": "voice-mcp",
"env": {
"WALLET_PRIVATE_KEY": "0x..."
}
}
}
}Optional:
{
"X402_VOICE_BASE_URL": "https://voice.forgemesh.io",
"BASE_RPC_URL": "https://mainnet.base.org"
}Notes
Paid tools require a Base wallet private key with USDC.
The server returns
audio_base64for audio tools so MCP clients can store, play, or forward the WAV bytes.No API keys or subscriptions are required for the voice service itself.
Maintenance
Latest Blog Posts
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/forgemeshlabs/voice-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server