mcp-speak
Related Servers
Alternatives to mcp-speak
AlicenseAqualityBmaintenanceConnects Speak AI transcription and insight data to Claude and ChatGPT, enabling natural language queries for summaries, action items, and quotes from recordings.100219 npmMIT- AlicenseNot gradedqualityCmaintenanceEnables AI assistants to speak aloud by generating and playing audio through the system output. Supports multiple TTS providers, playback queue management, and configurable voice profiles.11 npm1MIT
Related Servers
- AlicenseAqualityDmaintenanceEnables AI agents to synthesize natural speech using either platform system voices or premium OpenAI TTS, with automatic engine selection and graceful fallback.19 npmMIT
- FlicenseNot gradedqualityDmaintenanceEnables AI agents to generate and play high-quality text-to-speech audio using the Kokoro model, with support for multiple voices, adjustable speaking speed, and audio caching.-
- FlicenseNot gradedqualityCmaintenanceEnables coding agents to speak aloud using text-to-speech functionality. Works with agents running inside devcontainers and provides configurable voice settings for creating chatty AI companions.4-
- AlicenseNot gradedqualityDmaintenanceEnables agents to convert text to speech using OpenAI's TTS models with voice selection, delivery instructions, and queue-based audio playback. Supports both blocking and non-blocking modes for flexible audio generation and playback control.3BSD 3-Clause
- AlicenseAqualityDmaintenanceEnables AI agents to manage macOS audio routing, device switching, volume control, and multi-zone playback.16MIT
- FlicenseNot gradedqualityBmaintenanceEnables generating spoken audio from text and retrieving available voice characters and styles through MCP tools on macOS.-
TDQS
Scored across 2 tools
The two tools perform the same core action but are clearly differentiated by blocking behavior, with descriptive names and explicit notes in the descriptions. There is some overlap in purpose, but the distinction is obvious enough for an agent to select correctly.
Both tool names follow a consistent verb-first pattern with a modifier suffix (speak, speak_non_blocking). Naming clearly conveys the difference without mixing conventions or vague verbs.
Two tools is well-scoped for a text-to-speech server with blocking and non-blocking variants. Every tool earns its place, and the count matches the narrow purpose.
The server provides both blocking and non-blocking speech synthesis, covering the primary use cases. Minor gaps like canceling playback or selecting voices exist, but they are optional for the core domain.