"Using Google Search for Answer Generation" matching MCP connectors:
Matching Connector Tools:
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
Your AI rings your iPhone, speaks its question, and gets your spoken answer back as text.
Construction daily-log generation, jurisdiction compliance requirements, construction FAQs.
Human-input bridge for AI agents with voice-first answer links, MCP tools, and HTTP APIs.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
YouTube video search with transcript extraction as first-class output.
The ask-and-answer video format: send a question, get a short one-take video answer, captioned.
Search recordings, summarize meetings, create clips, and automate workflows from your AI assistant.
AI voice generation: text-to-speech and voice cloning from any MCP client.
Official MCP server for OmniDimension. Drive voice agents, dispatch calls, and run bulk campaigns.
Screen recording & video platform: search, share, transcribe, translate videos & AI meeting notes
Pronunciation scoring, speech-to-text, and text-to-speech for language learning
An MCP server that fetches video transcripts/subtitles, with pagination for large responses. Supports YouTube, Twitter/X, Instagram, TikTok, Twitch, Vimeo, Facebook, Bilibili, VK, Dailymotion, Reddit. Whisper fallback — transcribes audio when subtitles are unavailable.
Feedback layer for video. Reviewers talk through feedback; agents read it as structured comments.
Transcribe audio & video to text for AI agents: 100+ languages, speaker labels, webhooks.
Noogat is a voice-first note-taking app for iOS and web with an MCP server for AI coding agents. Capture ideas hands-free via Siri, retrieve them inside Claude Code, Claude Desktop, or Cursor. Features: semantic search, AI auto-tagging, search by time, related notes. Pro subscription required for MCP access.
Voice notes that organize themselves. Capture by Siri, AI auto-tags, semantic search retrieves.
Podcast intelligence for agents: transcripts, clips, speaker diarization, mention tracking.