Interactive Voice MCP Server
インタラクティブ音声 MCP サーバー (Kokoro TTS + NeMo ASR)
Kokoro を使用したテキスト読み上げ (TTS) 機能と、NVIDIA NeMo Parakeet モデルを使用した音声テキスト変換 (STT) 機能を提供し、対話型音声ダイアログを可能にするモデル コンテキスト プロトコル サーバーです。
利用可能なツール
interactive_voice_dialog- テキストを音声に合成して再生し、ユーザーの音声入力を聞いて書き起こしを返します。必要な引数:
text_to_speak(文字列): アシスタントが話すテキスト。
オプションの引数:
voice(文字列): TTSで使用する音声(例:'af_heart')。デフォルトは'af_heart'です。
インストール
前提条件
基礎となる TTS モデルの一部では、システムにespeak-ngがインストールされている必要があります。
Windows インストール:
espeak-ng リリースに移動します。
「最新リリース」をクリックします。
適切な
*.msiファイル (例:espeak-ng-20191129-b702b03-x64.msi) をダウンロードします。ダウンロードしたインストーラーを実行します。
ローカル開発インストール
Claude Desktop がpython -m mcp_server_ttsを使用してこのサーバーを起動できるようにするには、Python モジュールとしてインストールする必要があります。開発環境では、「編集可能」モード ( -e ) でインストールすることをお勧めします。これにより、ソースコードへの変更が再インストールなしで即座に反映されます。
pyproject.tomlファイル (このサーバー プロジェクトのルート) を含むディレクトリに移動し、次を実行します。
pip install -e .インストール後、次のコマンドを使用してスクリプトとして実行できます。
python -m mcp_server_tts.server # Assuming the main module is still server.py within mcp_server_tts
# Or, if you create a new package structure:
# python -m mcp_interactive_voice_serverRelated MCP server: Voice MCP
構成
Claude Desktopでこのサーバーを使用するには、 claude_desktop_config.jsonファイルに追加する必要があります。このファイルの場所は通常、 C:\Users\<YourUsername>\AppData\Roaming\Claude\claude_desktop_config.jsonです。
claude_desktop_config.jsonのmcpServersオブジェクトの下に次のエントリを追加します。
"tts": {
"command": "python",
"args": ["-m", "mcp_server_tts"]
}たとえば、 mcpServersセクションは次のようになります。
{
// ... other configurations ...
"mcpServers": {
// ... other servers ...
"tts": {
"command": "python",
"args": ["-m", "mcp_server_tts"]
}
// ... other servers ...
}
// ... other configurations ...
}This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables hands-free voice conversations with Claude using real-time speech recognition and text-to-speech on macOS. Creates a self-sustaining conversation loop where Claude can autonomously listen, respond, and continue the interaction without keyboard input.MIT
- AlicenseNot gradedqualityDmaintenanceEnables bidirectional voice interaction for Claude Code using local speech-to-text and text-to-speech models optimized for Apple Silicon. It provides tools to listen to user speech via microphone and speak responses aloud through system speakers.16Apache 2.0
- AlicenseNot gradedqualityDmaintenanceProvides text-to-speech generation using the Kokoro-82M model, enabling AI assistants to generate voiceovers and audio content directly within Claude Desktop and Cursor.14Apache 2.0
- AlicenseNot gradedqualityAmaintenanceEnables text-to-speech, voice cloning, audio generation, and transcription using Kokoro TTS and Whisper STT.131MIT
Related MCP Connectors
Voice and chat for AI agents — Discord, Teams, Meet, Slack, Zoom, Telegram, WhatsApp, NC Talk, SIP
Generate AI images, video, speech, music and presentations from Claude, ChatGPT and Cursor.
Connect Claude to Fathom meeting recordings, transcripts, and summaries
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/rungee84/voice_mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server