Skip to main content
Glama
569,274 tools. Updated 2026-09-14 22:20

"Text-to-speech tools for turning stories into podcasts" matching MCP tools:

  • Convert text to speech audio files using specified voices and models, saving results to your chosen directory for accessibility or content creation.
    MIT
  • Create custom voice profiles from audio samples for text-to-speech and speech-to-speech applications. Analyze MP3 or WAV files to generate voice replicas that mimic original audio characteristics.
    MIT
  • Convert text or notes files into natural MP3 speech. Submit content, get a job ID, and use it to check status for the finished audio.
    MIT
  • Create a new AI voice agent by configuring the language model, text-to-speech, and speech-to-text settings. Define the system prompt and voice to deploy a custom virtual assistant.
    MIT

Matching MCP Servers

Matching MCP Connectors

  • Generate highly realistic Text to Speech voiceovers.

  • An MCP server that provides tools to discover and retrieve podcast episodes transcripts.

  • Convert text into speech using ElevenLabs, returning an MP3 URL for playback. Specify a model and voice to customize the output; billed per 1,000 characters.
    MIT
  • Generate expressive WAV speech using one of 10 persona voices—like Storyteller or Announcer—for branded agents, characters, stories, and alerts.
    MIT
  • List the text-to-speech voice presets available for generating voice messages in email and SMS sent by AI agents.
    MIT
  • Convert text to speech and deliver it as a voice call to one or more phone numbers. Send automated voice messages directly to recipients.
    MIT
  • Check local speech-to-text backends and media tools, and get install instructions for any missing. Run before transcription to surface missing dependencies as clear answers.
    Apache 2.0
  • Transcribe any video or audio URL (YouTube, TikTok, podcasts) into timestamped text. Server-side processing handles download and speech-to-text; submit URL, get a job ID, poll for the transcript.
    MIT
  • Create a Text DAT in TouchDesigner linked to a vault note, loading its text into TD and syncing live with Obsidian edits for your visuals.
    MIT
  • Generate spoken audio from text using TTS synthesis. Supports multiple voices and audio formats for voiceovers, narrations, and podcasts.
    MIT
  • Turn written text into spoken audio using ElevenLabs voices. Choose fast or high-quality models, set speed and language, and save the speech file directly to your system.
    MIT