"Local RAG system for providing documentation to a large language model (LLM)" matching MCP connectors:
Matching Connector Tools:
Generate highly realistic Text to Speech voiceovers.
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
Remote MCP for AI video, image, music and speech generation.
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
ElevenLabs in natural language: generate speech in any language, create and manage voices, compose m
Any video URL to LLM-ready transcript. ASR built in, no captions needed. TikTok, X, TED and more.
Transcribe audio and video with Speechmatics speech-to-text from Claude and any MCP client.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Kurdish (Sorani & Kurmanji) text-to-speech & speech-to-text — 664 AI voices. API key required.
One key, 100+ models — chat with any LLM and generate video, images, speech. Free trial at 370.ai.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
Official MCP server for OmniDimension. Drive voice agents, dispatch calls, and run bulk campaigns.
Bambara AI over MCP: text-to-speech, transcription and translation (Bamanankan + more).
Your Plaud recordings in natural language: list recordings, read speaker-attributed transcripts and
Transcribe audio and video into speaker-labelled transcripts, subtitles, clips, and cited Q&A.
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning)
The ask-and-answer video format: send a question, get a short one-take video answer, captioned.
Outbound phone calls placed by an AI voice agent to book, ask, or confirm on your behalf.
AI voice generation: text-to-speech and voice cloning from any MCP client.