"How to Connect to a Remote Server" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
Media intelligence analysis for audio, video, and images via the Echosaw MCP server.
Generate AI images, video, voiceovers and music from Claude, ChatGPT or Cursor through 50+ models (Veo 3.1, Kling 3, Seedance, Nano Banana, GPT Image, ElevenLabs). Also image editing, upscaling, background removal, face swap, transcription, voice cloning and UGC-style video ads. Sign in with OAuth — no API key to paste. Tools are annotated (read-only vs. credit-spending); failed generations are refunded.
Create and edit images, videos, and audio through Magic Hour's hosted Streamable HTTP MCP server.
Change lyrics in an existing song. Agents upload authorized MP3 audio, provide a 0.1–6 second phrase and replacement words, preview generated singing, and export MP3/WAV. No website signup required: buy credits via x402 with USDC, then use a scoped key. Caller supplies timing. Resumable CLI and Codex/Claude Code setup: https://lyricpatch.com/agents/integrations. Hear a real example: https://lyricpatch.com/agents/demo.
One tool surface for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one credit pool. Connect in one click with OAuth, no API key required.
Deepfake detection, media intelligence, and invisible watermarking for audio, image, and video via the Resemble AI API, plus docs tools. Remote MCP server (Streamable HTTP) — also published in the official MCP registry as io.github.resemble-ai/resemble-mcp.
Process video, audio, images, and documents with 86+ cloud media processing robots.
Human-made production music for sync — search by brief or reference, preview, score to picture.
MCP server for RiverScript, an AI transcription platform - fetches transcripts shared via a link.
AI transcription from URLs or files. 119 languages, diarization, SRT/VTT/text export.
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
Remote MCP for AI video, image, music and speech generation.
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Decode an SSTV audio recording and anchor its fingerprint and image hash to the Knox event chain.
Detect AI-generated images, videos, and audio with identifAI's deepfake detection tools.
AI audio tools for music producers — stem splitting, vocal removal, BPM & key detection, audio-to-MIDI, format conversion, trimming, video-to-audio extraction and AI song generation.
Arabic-first AI creative platform for Egyptian and Arab businesses. Generate social media designs, write marketing copy in Egyptian dialect, build content calendars, produce Sora-2 videos, AI photoshoots, music tracks, and business documents — with your brand identity automatically applied. Requires a Grow or Business subscription at vizzy.space.
25+ AI media generation tools — FLUX Pro, Ideogram v3, Recraft v3, Stable Diffusion XL, MiniMax video, and Kokoro TTS. Images, video, and audio from one server. $0.01/call.
Focused MCP server for OpenAI image/audio generation (v2.0.0). Wraps endpoints via HAPI CLI.
Generate highly realistic Text to Speech voiceovers.