"Creating custom experimental images in Docker" matching MCP connectors:
Matching Connector Tools:
Media intelligence analysis for audio, video, and images via the Echosaw MCP server.
Generate marketing images, videos and audios for campaigns, product content, and brand assets.
One tool surface for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one credit pool. Connect in one click with OAuth, no API key required.
Deepfake detection, media intelligence, and invisible watermarking for audio, image, and video via the Resemble AI API, plus docs tools. Remote MCP server (Streamable HTTP) — also published in the official MCP registry as io.github.resemble-ai/resemble-mcp.
Generate and edit images, videos, and audio with 150+ models from 20+ vendors.
Train portable RVC v2 voice models from audio in the cloud and download the .pth, .index, and ZIP.
Process video, audio, images, and documents with 86+ cloud media processing robots.
LibriVox public-domain audiobooks (~17000 titles in dozens of languages)
Arabic-first AI creative platform for Egyptian and Arab businesses. Generate social media designs, write marketing copy in Egyptian dialect, build content calendars, produce Sora-2 videos, AI photoshoots, music tracks, and business documents — with your brand identity automatically applied. Requires a Grow or Business subscription at vizzy.space.
Detect AI-generated images, videos, and audio with identifAI's deepfake detection tools.
AudioAlpha turns 100+ daily finance and crypto podcasts into structured intelligence — α-sentiment scores, narrative signals, asset mentions, transcripts, and market snapshots with 40+ custom metrics. Built for AI-driven research and trading workflows.
25+ AI media generation tools — FLUX Pro, Ideogram v3, Recraft v3, Stable Diffusion XL, MiniMax video, and Kokoro TTS. Images, video, and audio from one server. $0.01/call.
Generate images, videos, voiceovers, and captions from a chat prompt.
Generate images, video, and audio with Glif's media-generation agent
Search speech in podcasts, government meetings, and your own audio: speakers, entities, timestamps.
Generate images, video, music, voice and 3D through one API. 30 tools, 200+ models.
Turn Claude into a creative studio: DNA-locked characters, images, video, voiceover — 55 tools.
Turn any LLM multimodal; generate images, voices, videos, 3D models, music, and more.
Send web articles or AI-written text to your personal podcast feed; listen in any podcast app.