"Methods for Reading PDF Files" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
One tool surface for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one credit pool. Connect in one click with OAuth, no API key required.
CC0 sound effects API for AI agents — search, preview, and download via MCP.
Audio for your agent: transcribe, speak, translate, summarise, plus sound effects and music.
Media intelligence analysis for audio, video, and images via the Echosaw MCP server.
Deepfake detection, media intelligence, and invisible watermarking for audio, image, and video via the Resemble AI API, plus docs tools. Remote MCP server (Streamable HTTP) — also published in the official MCP registry as io.github.resemble-ai/resemble-mcp.
Generate marketing images, videos and audios for campaigns, product content, and brand assets.
Deterministic music theory for agents: analyze, voice, reharmonize, conduct — computed, not guessed
CC0 sound effects API for AI agents — search, preview, and download via MCP.
Music superpowers for your AI agent. Curated licensed music and AI tools that do the rest. Right where you already work.
AI audio tools for music producers — stem splitting, vocal removal, BPM & key detection, audio-to-MIDI, format conversion, trimming, video-to-audio extraction and AI song generation.
Arabic-first AI creative platform for Egyptian and Arab businesses. Generate social media designs, write marketing copy in Egyptian dialect, build content calendars, produce Sora-2 videos, AI photoshoots, music tracks, and business documents — with your brand identity automatically applied. Requires a Grow or Business subscription at vizzy.space.
File conversion: PDF, DOCX, STT, TTS, watermarking
AudioAlpha turns 100+ daily finance and crypto podcasts into structured intelligence — α-sentiment scores, narrative signals, asset mentions, transcripts, and market snapshots with 40+ custom metrics. Built for AI-driven research and trading workflows.
Transcribe Russian audio and video with timestamps and optional speaker labels from files or links.
Free CC0 sound effects for agents: ask by role (button-click, coin), sets, or search 4,600+.
Human-made production music for sync — search by brief or reference, preview, score to picture.
MCP server for RiverScript, an AI transcription platform - fetches transcripts shared via a link.
One API for 100+ AI video, image, music and speech models.
AI transcription from URLs or files. 119 languages, diarization, SRT/VTT/text export.
Video, audio, and image processing for AI agents: convert, transcribe, upscale - 150+ operations.