ThotStream
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@ThotStreamSpeak this thought aloud: I should verify the database connection first."
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
ThotStream Lite 🎙️
Minimalist, zero-dependency, cross-compatible AI-to-audio speech engine and MCP server.
Stream internal thoughts, tool actions, and responses in real-time with 1 local voice.
⚡ Core Highlights
Zero Mandatory Dependencies: Runs entirely on Python 3.8+ standard library (
subprocess,threading,queue,shutil).Verified Cross-Platform:
macOS: Built-in
sayCLI andafplayCoreAudio playback (0 MB install).Linux: Intelligent auto-cascade across
spd-say(speech-dispatcher),espeak-ng,espeak, plus ALSA (aplay), PulseAudio (paplay), and PipeWire (pw-play).Windows: Built-in
System.Speechvia PowerShell (0 MB install).Any OS (Neural Upgrade): Automatically detects local Piper TTS ONNX models for neural voice output.
Universal Multi-Adapter:
MCP Server: Stdio JSON-RPC 2.0 compliant with Claude Desktop, Antigravity IDE, Cursor, and Continue.
CLI Pipe & Wrapper: Transparent
stdoutinterception (thotstream-wrap) for zero token overhead in terminal harnesses (Freebuff, Gemini CLI, Claude Code CLI).Skill Card: Standard
SKILL.mdinstruction specification for prompt-driven agents.
Non-Blocking Threaded Architecture: Speech synthesis and audio playback occur on an asynchronous worker thread, ensuring LLM text generation is never blocked.
Related MCP server: mcp-ai-voice
🚀 Quickstart (60 Seconds)
1. Install (Editable / Zero Dependencies)
cd packages/thotstream-lite
pip install -e .2. Verify Your System Audio Driver
thotstream-lite --statusExample outputs:
macOS:
Audio Driver: sayLinux:
Audio Driver: spd-say(orespeak-ng)Windows:
Audio Driver: sapi5With Piper:
Audio Driver: piper
🔌 Integration Modes
Mode A: Claude Desktop & Antigravity IDE (MCP)
Add to claude_desktop_config.json or .gemini/settings.json:
{
"mcpServers": {
"thotstream": {
"command": "python",
"args": ["-m", "thotstream_lite.mcp_server"]
}
}
}The agent receives three dedicated audio tools:
speak_thought(text): Narrate internal reasoning or hypotheses.speak_action(text): Announce tool execution intent before running.speak_response(text): Speak the final answer aloud.
See docs/INTEGRATIONS.md for full setup screenshots.
Mode B: CLI Streaming Interception (Freebuff / Terminal CLIs)
Zero token overhead. The LLM generates text normally with XML tags; ThotStream Lite intercepts and speaks them out-of-band:
# 1. Pipe streaming stdout from any agent
freebuff --mode agent | thotstream-lite --listen
# 2. Or wrap the CLI command directly
thotstream-lite freebuff --mode agentMode C: Python SDK
from thotstream_lite import ThotStreamLite
engine = ThotStreamLite()
# Asynchronous, non-blocking queue calls
engine.speak_thought("Formulating system architecture hypothesis.")
engine.speak_action("Querying vector database for matching nodes.")
engine.speak_response("Operation completed successfully.")
# Drain and shutdown
engine.drain()
engine.shutdown()📂 Documentation & Examples
📘 Architecture Deep-Dive: Threaded queue model, driver dispatch, and fail-open guarantees.
🌐 Platform Compatibility Matrix: Complete Linux, macOS, and Windows compatibility details.
🔌 Integration Guides: Config guides for Claude Desktop, Antigravity, and Freebuff.
⚖️ Legal Disclaimer & Responsible AI Use
THOTSTREAM LITE IS DISTRIBUTED UNDER THE MIT LICENSE ON AN "AS IS" AND "AS AVAILABLE" BASIS, WITHOUT WARRANTIES OR GUARANTEES OF ANY KIND.
User Responsibility: The user/operator assumes 100% legal and operational responsibility for any text ingested, commands executed, and audio synthesized through this software.
Voice Rights & Regulations: Maintainers do not bundle or license third-party voice models. Users are solely responsible for compliance with voice likeness, copyright, and AI synthesis regulations.
Read the full legal notice in DISCLAIMER.md and LICENSE.
This server cannot be deployed
Maintenance
Related MCP Connectors
Audio for your agent: transcribe, speak, translate, summarise, plus sound effects and music.
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
AI voice generation: text-to-speech and voice cloning from any MCP client.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
Related MCP Servers
- FlicenseNot gradedqualityAmaintenanceEnables coding agents to speak aloud using text-to-speech functionality. Works with agents running inside devcontainers and provides configurable voice settings for creating chatty AI companions.6-
- AlicenseAqualityDmaintenanceEnables AI agents to synthesize natural speech using either platform system voices or premium OpenAI TTS, with automatic engine selection and graceful fallback.114 npmMIT
- AlicenseNot gradedqualityCmaintenanceEnables AI assistants to speak aloud by generating and playing audio through the system output. Supports multiple TTS providers, playback queue management, and configurable voice profiles.32 npm1MIT
- AlicenseNot gradedqualityBmaintenanceEnables voice-first interactions with AI agents and MCP tools, supporting speech input/output, STT/TTS, and a provider-independent agent core.1MIT