"A server for making LLMs follow user prompts more strongly" matching MCP connectors:
Matching Connector Tools:
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
Any video URL to LLM-ready transcript. ASR built in, no captions needed. TikTok, X, TED and more.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Voice-to-PDF daily construction reports. Court-ready documentation for subcontractors
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
YouTube video search with transcript extraction as first-class output.
The ask-and-answer video format: send a question, get a short one-take video answer, captioned.
Search recordings, summarize meetings, create clips, and automate workflows from your AI assistant.
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning)
Official MCP server for OmniDimension. Drive voice agents, dispatch calls, and run bulk campaigns.
User research workspace to transcribe interviews and turn conversations into insights.
Voice-interview your ideas into LinkedIn posts and X threads via a Digital Brain memory.
Give your AI agent a phone: place calls, navigate IVRs, wait on hold, get structured answers.
Pronunciation scoring, speech-to-text, and text-to-speech for language learning
Feedback layer for video. Reviewers talk through feedback; agents read it as structured comments.
An MCP server that fetches video transcripts/subtitles, with pagination for large responses. Supports YouTube, Twitter/X, Instagram, TikTok, Twitch, Vimeo, Facebook, Bilibili, VK, Dailymotion, Reddit. Whisper fallback — transcribes audio when subtitles are unavailable.
Transcribe audio & video: diarization, timed SRT/VTT, podcasts, paste-a-link, whole-feed batch.
Give AI agents a phone: outbound AI calls that return a summary, transcript, and extracted fields.