m4xSpeech ProcessingAudio ProcessingcloseofbusinessAlicenseAqualityDmaintenanceLocal, private audio transcription MCP server enabling AI agents to transcribe audio files entirely on-device without uploading data. Updated 3 months ago (2026-07-10 20:34 UTC)3MIT
this-needs-a-callSpeech ProcessingAutonomous AgentsjacobparisFlicense-Not gradedqualityDmaintenanceSelf-host a realtime voice-call companion for coding agents. Exposes an MCP endpoint that agents can poll as an alternate input stream. Updated a month ago (2026-08-20 02:14 UTC)18-
io.github.chicogong/ffvoiceSpeech ProcessingAudio ProcessingchicogongAlicense-Not gradedqualityBmaintenanceMCP server for offline speech-to-text and speaker diarization, enabling AI agents to transcribe audio locally without cloud APIs. Updated 18 days ago (2026-09-15 04:35 UTC)104 PyPI3MIT
ollos-mcpMultimedia ProcessingImage & Video ProcessingAudio ProcessingkelvinbiffiAlicense-Not gradedqualityAmaintenanceProvides local, offline transcription, keyframe extraction, OCR, and pre-publish review of audio, video, and image files, enabling AI agents to see and hear media without cloud or API keys. Updated 11 days ago (2026-09-22 02:34 UTC)36 npmApache 2.0
STT2TTS MCPSpeech ProcessingText-to-SpeechpygodzillaAlicense-Not gradedqualityDmaintenanceLocal-first speech-to-text and text-to-speech MCP server. Hot-swappable engines via config.yaml — no code changes, no API keys required. Updated 3 months ago (2026-06-19 18:42 UTC)2MIT
voice-mcp-serverSpeech ProcessingAudio ProcessingerickvsAlicense-Not gradedqualityCmaintenanceEnables AI agents to speak and listen in real-time with interruption handling, using local ML models and hot-swappable adapters. Updated 3 months ago (2026-06-28 18:04 UTC)32 npmMIT