MCP server that enables natural language control of internet radio from Claude Code, with access to 30,000+ global stations, auto-playback via mpv, and a real-time status line with audio spectrum visualization.
Enables converting text to speech audio in 20+ languages, returning base64-encoded MP3 output via Google TTS with x402 micropayment-based pay-per-call access.
This MCP server enables AI tools to read Dedao Brain (formerly Get Notes) web notes from a free account, exposing AI summaries, original transcripts, and audio attachment links or downloads, with an accompanying CLI.
Text-to-speech MCP server that enables AI assistants to read text aloud on the user's computer using Windows SAPI, with no API key or cloud service required.
Enables conversion of YouTube videos to MP3 format through the Youtube To Mp315 API. Supports checking conversion status, retrieving video titles, and asynchronous video-to-audio conversion with customizable quality and time range settings.
Enables downloading videos from platforms like YouTube and converting them to text using OpenAI Whisper and ffmpeg. It supports multiple output formats including TXT, JSON, SRT, and VTT for transcriptions.
Enables users to search movies and actor profiles, discover soundtracks, and generate curated Spotify playlists, with full playback control and playlist management through natural language.
Enables text-to-speech synthesis using Coqui TTS with multiple models and languages, long-text chunking, customizable speed and speaker settings, and voice cloning from reference audio.
Enables users to convert text into high-quality audio by accessing the OpenAI Text-to-Speech API. It supports customizable model selection and voice options for synthesized speech generation via the MCP protocol.
An MCP server that enables voice-to-voice AI conversations using ElevenLabs for speech synthesis and recognition, with tools for voice management, text-to-speech, and speech-to-text.
Enables submitting audio via URL or Base64 for local Whisper transcription, then checking job status and results or deleting transcription jobs. It provides asynchronous job handling and SSRF-protected downloads/callbacks without cloud API calls.
Enables deterministic room acoustics analysis through MCP tools that compute axial, tangential, and oblique standing wave modes, Bonello and Bolt criteria compliance, Schroeder cutoff frequency, and Sabine/Norris-Eyring RT60 decay across octave bands for rectangular rooms. Also derives optimal speaker and sweet-spot coordinates with SBIR notch predictions and calculates the acoustic treatment area needed to hit target reverb times for podcast, mixing, home theater, or listening use.
Enables AI-powered music generation and live coding by providing direct control over Strudel.cc through browser automation. Supports pattern creation, audio analysis, and pattern storage for TidalCycles/Strudel music patterns.