"A resource for academic research or scholarly articles" matching MCP connectors:
Matching Connector Tools:
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Voice-to-PDF daily construction reports. Court-ready documentation for subcontractors
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
The ask-and-answer video format: send a question, get a short one-take video answer, captioned.
Outbound phone calls placed by an AI voice agent to book, ask, or confirm on your behalf.
AI voice agents that make real phone calls: single calls or campaigns, with transcripts and notes.
Record your pitch in Claude or ChatGPT and get instant feedback on your delivery.
Official MCP server for OmniDimension. Drive voice agents, dispatch calls, and run bulk campaigns.
User research workspace to transcribe interviews and turn conversations into insights.
Voice-interview your ideas into LinkedIn posts and X threads via a Digital Brain memory.
Give your AI agent a phone: place calls, navigate IVRs, wait on hold, get structured answers.
Pronunciation scoring, speech-to-text, and text-to-speech for language learning
Feedback layer for video. Reviewers talk through feedback; agents read it as structured comments.
An MCP server that fetches video transcripts/subtitles, with pagination for large responses. Supports YouTube, Twitter/X, Instagram, TikTok, Twitch, Vimeo, Facebook, Bilibili, VK, Dailymotion, Reddit. Whisper fallback — transcribes audio when subtitles are unavailable.
Transcribe audio & video: diarization, timed SRT/VTT, podcasts, paste-a-link, whole-feed batch.
Give AI agents a phone: outbound AI calls that return a summary, transcript, and extracted fields.
Transcribe audio & video to text for AI agents: 100+ languages, speaker labels, webhooks.