Enables text-to-speech and speech-to-text through MCP tools, using Groq's free hosted endpoints when configured and falling back to fully local keyless models otherwise. Also provides voice listing and provider health checks.
Enables spoken conversations with Claude on a Mac: Claude speaks through speakers, listens to the user's natural replies, and transcribes them locally. No audio leaves the computer, and it includes tools for voice setup and a hands-free voice mode.
An MCP server that converts short text summaries into spoken audio and plays them locally, designed for coding agents like Claude Code to announce task results.
Voice interface for Claude Code: you talk, the agent listens, codes, and talks back while it works. Live speech-to-text with turn-taking, Grok/xAI voices with per-subagent personas, and a real-time HUD dashboard.
Enables any MCP-compatible client to perform local voice cloning and text-to-speech using OmniVoice, including voice profile management, voice design, and an optional ASMR post-processing DSP pipeline.
Enables AI agents to present interactive code walkthroughs with voice narration, opening files, highlighting code, and showing inline explanations with synchronized text-to-speech.
Official Model Context Protocol (MCP) Server for QuickerSpot — AI-powered commercial radio and retail sound automation.
Connect your AI Assistants (Cursor IDE, Claude Desktop, Antigravity, Hermes Agent, OpenClaw) directly to QuickerSpot's ElevenLabs V3 voice engine, AI script generator, and indoor radio mixer.
Enables agents to speak Thai and English audio through local speakers, manage a playback queue, save speech to files, translate text into Thai, and OCR PDFs and images.
A language expression coach MCP server that provides native-sounding translations with cultural context, tone notes, and audio pronunciation for over 13 languages.
Enables local text-to-speech synthesis for Claude and Cursor using Supertonic 3, with support for multiple voices, expressions, and languages. No API key or cloud required.
MCP server for the Supertone TTS API. Generate natural speech, browse and preview the
voice catalog, predict synthesis cost, and create cloned voices — directly from Claude
Desktop, Cursor, or any MCP-compatible client. Supports Korean, English, Japanese, and
20+ other languages, with speed, pitch, and emotion-style control.
Image, video, chat and text-to-speech models (GPT Image, Gemini Image, Veo, Kling, Seedance, Claude, GPT, Gemini) behind one API key. generate_image returns a preview the model can see, and review_image has a vision model critique the result and propose a corrected prompt, so an assistant can generate, check and fix images on its own.