"Tools or methods to automate browser interactions with clicks" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
Generate highly realistic Text to Speech voiceovers.
Bambara AI over MCP: text-to-speech, transcription and translation (Bamanankan + more).
AI transcription from URLs or files. 119 languages, diarization, SRT/VTT/text export.
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
Voice-powered bug reporting with 13 MCP tools. Record bugs by talking; let AI find and fix them.
Any video URL to LLM-ready transcript. ASR built in, no captions needed. TikTok, X, TED and more.
Transcribe audio and video with Speechmatics speech-to-text from Claude and any MCP client.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Kurdish (Sorani & Kurmanji) text-to-speech & speech-to-text — 664 AI voices. API key required.
One key, 100+ models — chat with any LLM and generate video, images, speech. Free trial at 370.ai.
AI voice agents on SMB websites — fully autonomous build in 2–3 min. 23 MCP tools. EU, GDPR.
YouTube video search with transcript extraction as first-class output.
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
Search recordings, summarize meetings, create clips, and automate workflows from your AI assistant.
MCP server exposing the AceDataCloud Fish Audio API (text-to-speech with voice conditioning)
Hosted MCP for speech cleanup. Remove noise, filler words, and dead air. Also cut, join, transcribe, and TTS. Sign in with a CleanAudio account. No API key. Endpoint: https://mcp.cleanaudio.app/mcp This is cleanaudio.app, not cleanaudio.io.
Pronunciation scoring, speech-to-text, and text-to-speech for language learning
AI speech-to-text for public TikTok videos: SRT, VTT, word timings, speaker labels, 90+ languages.
An MCP server that fetches video transcripts/subtitles, with pagination for large responses. Supports YouTube, Twitter/X, Instagram, TikTok, Twitch, Vimeo, Facebook, Bilibili, VK, Dailymotion, Reddit. Whisper fallback — transcribes audio when subtitles are unavailable.