"How to configure Azure Speech" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Start with classify_media_route to route long or oversized audio/video by size or duration.
Podcast post-production for an AI agent: transcribe a recording, cut fillers and pauses, clean up the voice, add licensed music that ducks under speech, set chapters and export a finished episode.
Search, read and reply to your Telegram Business chats, transcribed voice included.
Find and pay for AI inference per call (chat, embeddings, transcription, images) from GPU sellers and agentic.no's own pool. Pay with x402 (USDC on Base); no account or API key. Includes Norwegian speech-to-text (NB-Whisper).
AI voice generation: text-to-speech and voice cloning from any MCP client.
One key, 100+ models: chat with any LLM, generate images and speech. Live model list. Free trial.
Text to speech in 149 languages: MP3 links from any assistant. Free without an account.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Sorani & Kurmanji TTS+STT: Kurdish speech most APIs lack. 885 voices, free tier, no key to browse.
Pay-per-call AI gateway: models, speech, web search, on-chain reads and NLP tools, via x402.
Hosted speech-to-text + speech emotion/tone analysis for agents. No install; trial keys built in.
Bambara AI over MCP: text-to-speech, transcription and translation (Bamanankan + more).
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
Pronunciation assessment, phoneme scoring, speaker voice ID, audio transcription, speech synthesis.
Remote MCP for AI video, image, music and speech generation.
Any video URL to LLM-ready transcript. ASR built in, no captions needed. TikTok, X, TED and more.
ElevenLabs in natural language: generate speech in any language, create and manage voices, compose m
User research workspace to transcribe interviews and turn conversations into insights.
Create, inspect, and manage Wubble music, speech, voice, and sound-effect requests through MCP.
Get live AI suggestions for what to say in job interviews, sales calls, and meetings, tailored to your CV, job description, client brief, and notes.