EasyVoice text to speech
Server Details
Hosted text-to-speech MCP server: 66 voices in 9 languages, long-form narration, podcasts.
- Status
- Healthy
- Last Tested
- Transport
- Streamable HTTP · MCP 2025-11-25
- URL
TDQS
Score is being calculated.
Available Tools
7 toolsclone_voiceClone a VoiceInspect
Enroll a custom cloned voice from an audio sample (async). Requires a Pro or active 7-day pass key. Provide exactly one of audio_base64 or audio_url (https). Sample limits: 10 MB max upload; wav, mp3 or m4a; a 10–30 second clean single-speaker recording (enrollment rejects clips shorter than 10 or longer than 30 seconds; trim longer recordings first). Account limits: maximum 3 cloned voices, 5 enrollments per day. Returns a voice_id with status "enrolling" — check list_voices until its status is "ready", then use it in create_podcast.
| Name | Required | Description | Default |
|---|---|---|---|
| name | Yes | Display name for the cloned voice. | |
| consent | Yes | Required attestation: you confirm you have the recording rights to this audio sample. The server rejects anything but an explicit true. | |
| audio_url | No | https URL to the audio sample — fetched server-side, same 10 MB bound. | |
| mime_type | No | Sample MIME type — wav, mp3 and m4a are accepted (the server enforces the exact allowlist). | audio/wav |
| audio_base64 | No | Base64-encoded audio sample (10 MB max decoded). |
Output Schema
| Name | Required | Description |
|---|---|---|
| status | Yes | |
| voice_id | Yes |
create_podcastCreate Podcast EpisodeInspect
Create a two-host (A/B) podcast episode as an async job. Requires a Pro subscription key. Up to 60 segments; combined text up to 30,000 characters. voice_a and voice_b set the two speaker voices — your READY cloned voice_* ids are allowed. Returns a job id — poll get_job_status.
| Name | Required | Description | Default |
|---|---|---|---|
| format | No | mp3 | |
| voice_a | Yes | Voice id for speaker A — catalog voice or one of your READY cloned voice_* ids. | |
| voice_b | Yes | Voice id for speaker B. | |
| segments | Yes | Alternating dialogue segments (max 60). |
Output Schema
| Name | Required | Description |
|---|---|---|
| job_id | Yes | |
| status | Yes |
generate_long_formGenerate Long-Form NarrationInspect
Generate long-form narration (up to 500,000 characters ≈ 9 hours of audio) as an async job. Requires a Pro subscription key. Kokoro narration voices only — Arabic (ar_*) and cloned (voice_*) voices are rejected by the Long-Form engine. Returns a job id — poll get_job_status; parts render progressively and audio_url appears when completed.
| Name | Required | Description | Default |
|---|---|---|---|
| tone | No | Tone preset. | neutral |
| input | Yes | The full script, up to 500,000 characters. | |
| pitch | No | Pitch shift in semitones (-4 to 4). | |
| speed | No | ||
| voice | No | Kokoro voice id — call list_voices for the catalog. | af_aoede |
| format | No | mp3 |
Output Schema
| Name | Required | Description |
|---|---|---|
| job_id | Yes | |
| status | Yes | |
| input_chars | Yes |
get_job_statusGet Job StatusRead-onlyInspect
Check the status of a long-form, podcast or tts job owned by this key. Returns status plus audio URL(s) when completed. Audio links are not guaranteed to persist — download promptly.
| Name | Required | Description | Default |
|---|---|---|---|
| job_id | Yes | Job id returned by an async tool. |
Output Schema
| Name | Required | Description |
|---|---|---|
| id | Yes | |
| status | Yes | |
| audio_url | No |
get_usageGet UsageRead-onlyInspect
Check this API key's plan and usage. Free keys see today's used/remaining characters (shared 5,000/day pool, resets at midnight UTC), the 20 requests/minute rate limit and 2-concurrent cap. Pro keys are unlimited (fair use: 10M characters per rolling 30 days, then slowed to 30 requests and 10,000 characters a minute, 10 and 3,000 past 20M, on every surface -- never blocked or billed) at 60 requests/minute.
| Name | Required | Description | Default |
|---|---|---|---|
No parameters | |||
Output Schema
| Name | Required | Description |
|---|---|---|
| plan | Yes | |
| used | No | |
| limit | No | |
| reset_at | No | |
| remaining | No | |
| rpm_limit | Yes | |
| unlimited | No | |
| concurrency_limit | No | |
| fair_use_chars_per_month | No |
list_voicesList VoicesRead-onlyInspect
List all 66 EasyVoice catalog voices (id, name, language, accent, gender, free/pro tier) plus this key's cloned voices. Call this before text_to_speech or create_podcast to pick voice ids.
| Name | Required | Description | Default |
|---|---|---|---|
No parameters | |||
Output Schema
| Name | Required | Description |
|---|---|---|
| plan | Yes | |
| cloned | Yes | |
| voices | Yes |
text_to_speechText to SpeechInspect
Convert text to a downloadable MP3 or WAV audio file with one of EasyVoice's 66 neural voices. Returns a hosted audio URL (no inline audio). Per-call limit: 8,000 characters (4,000 for Arabic ar_* voices) — text over the limit is automatically queued as a background job instead of erroring; poll get_job_status with the returned job_id for the result. Free keys can use the 12 free voices; Pro ($9.99/mo) unlocks the full catalog. A Pro account past the fair-use line is paced in characters a minute: a call may wait up to 20 seconds, and one that would wait longer returns a tool error saying when to retry.
| Name | Required | Description | Default |
|---|---|---|---|
| text | Yes | The text to speak. | |
| speed | No | ||
| voice | No | Voice id — call list_voices for the catalog. | af_aoede |
| format | No | mp3 |
Output Schema
| Name | Required | Description |
|---|---|---|
| voice | Yes | |
| format | Yes | |
| audio_url | Yes | |
| characters | Yes | |
| remaining_free_chars_today | No |
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
7 tool updates
- First observed
clone_voice - First observed
create_podcast - First observed
generate_long_form - First observed
get_job_status - First observed
get_usage - First observed
list_voices - First observed
text_to_speech
Related MCP Connectors
MCP server for Text-to-Speech
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning)
MCP server for Speech-to-Text
Hosted pay-per-use TTS: 54 neural voices, 9 languages incl. Brazilian Portuguese. $10 free credits.
Related MCP Servers
- AlicenseNot gradedqualityAmaintenanceHosted text-to-speech MCP server for AI agents with 54 neural voices in 9 languages, including Brazilian Portuguese. Pay-per-use API, no GPU or subscriptions needed.MIT
- FlicenseAqualityDmaintenanceMCP server for text-to-speech synthesis using Azure Speech Services, supporting 6 languages with high-quality neural voices and smart voice selection.1-
- AlicenseNot gradedqualityDmaintenanceAn MCP server that converts text into lifelike speech using Microsoft Edge's Text-to-Speech service, supporting customizable voice, rate, volume, and pitch.4MIT
- AlicenseAqualityBmaintenanceText to speech in 149 languages: MP3 links from any assistant. Free without an account. 2,253 voices; PRO adds HD voices, WAV, dialogue with a voice per speaker, Script mode timing and transcription.762 npmMIT
Glama MCP Gateway
Add one secure layer between your agents and this server.