"Exploring text-to-image generation techniques" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
Convert images to PNG, JPEG, WebP, or AVIF through one public remote MCP tool.
The CNAPS.ai MCP Server lets Claude build and run AI image, video, and text pipelines mid-conversation — no dashboard, no manual node-wiring. Describe a task in plain language (e.g. "upscale this photo 4x") and Claude selects the right model(s), wires a pipeline, and runs it.
One tool surface for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one credit pool. Connect in one click with OAuth, no API key required.
Create long-form YouTube videos end to end: script, storyboard, voiceover, final MP4.
Extract tables, text and formulas from PDFs, including scanned pages and broken text layers.
Free CC0 sound effects for agents: search 4,000+ sounds, hotlink MP3/WAV, generate SFX from text.
Give AI random access to video: timestamped contact sheets + zoom into any start/end range.
Create narrated presentations from HTML, poll build status, list them, read one back as text.
A one-stop creative pipeline for AI agents: generate, upscale, enrich, sign, store, mint. 24 paid MCP tools powered by Stable Diffusion, Imagen 3, ESRGAN, and Gemini — plus 53K+ museum artworks from Alexandria Aeternum. Three payment rails, volume discounts, and a free trial to start.
Turn one line-art image into a prompt-directed, hand-drawn scribe animation MP4.
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
Render HTML, Markdown, or URLs to images, PDF, or branded artifacts; extract and watch pages.
Video, audio, and image processing for AI agents: convert, transcribe, upscale - 150+ operations.
Screenshot, PDF, OG-image, and page extraction (markdown/JSON) over MCP. Bearer key or x402.
Image processing for AI agents: resize, convert, compress, crop, and web-ready AI-generated images.
Any video URL to LLM-ready transcript. ASR built in, no captions needed. TikTok, X, TED and more.
Run multi-step AI pipelines for video, image, audio and text: upload media, run, poll results.
Compress images to target sizes (80%+ smaller). Tools: compress_image, convert_image, optimize_image
Transcribe audio and video with Speechmatics speech-to-text from Claude and any MCP client.
Zero-auth MCP: image optimize, cited storage/format data, and dev utilities LLMs get wrong.