"How to run a Windows command" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
Generate highly realistic Text to Speech voiceovers.
Generate game assets with AI: sprites, 3D models, animations, sound effects, music, and voices.
Deepfake detection, media intelligence, and invisible watermarking for audio, image, and video via the Resemble AI API, plus docs tools. Remote MCP server (Streamable HTTP) — also published in the official MCP registry as io.github.resemble-ai/resemble-mcp.
Generate AI images, video, music, and sound effects, and upscale them, from any MCP client.
Human-made production music for sync — search by brief or reference, preview, score to picture.
MCP server for RiverScript, an AI transcription platform - fetches transcripts shared via a link.
Verbatim transcription of public video/audio URLs to clean text, SRT, and timestamped records.
LibriVox public-domain audiobooks (~17000 titles in dozens of languages)
Connect the OneStepTranscribe MCP server to your AI assistant and turn audio or video into text without leaving the chat. It is a remote server, so there is nothing to install, no API key, and no account. Just add one URL and ask your assistant to transcribe a file.
Decode an SSTV audio recording and anchor its fingerprint and image hash to the Knox event chain.
AI audio tools for music producers — stem splitting, vocal removal, BPM & key detection, audio-to-MIDI, format conversion, trimming, video-to-audio extraction and AI song generation.
Arabic-first AI creative platform for Egyptian and Arab businesses. Generate social media designs, write marketing copy in Egyptian dialect, build content calendars, produce Sora-2 videos, AI photoshoots, music tracks, and business documents — with your brand identity automatically applied. Requires a Grow or Business subscription at vizzy.space.
Focused MCP server for OpenAI image/audio generation (v2.0.0). Wraps endpoints via HAPI CLI.
125+ browser tools for PDF, Image, Video, Audio, AI, Scanner. Files never leave your device.
Hosted MCP server that gives AI agents (Claude, Cursor, Codex, etc.) access to the full Runware API — image generation, video generation, audio generation, 3D, upscaling, background removal, captioning, and more.
Give ears to Claude/Openclaw/Hermes/Codex/Grok Bot. Voibe turns recordings into text your AI agent can work with. Ask your agent to transcribe a meeting, interview, call, lecture, podcast episode or voice memo. The raw transcript arrives in the chat with speaker labels, timestamps and a summary. Attach the file in the chat, or point at a file or folder in Claude Code, where a whole folder of recordings works in
Generate and edit images, videos, and audio with 150+ models from 20+ vendors.
Your assistant curates. Rovyn plays. Podcast plans become tap-to-play editions; receipts come back.
Hosted MCP for speech cleanup. Remove noise, filler words, and dead air. Also cut, join, transcribe, and TTS. Sign in with a CleanAudio account. No API key. Endpoint: https://mcp.cleanaudio.app/mcp This is cleanaudio.app, not cleanaudio.io.
The Listenetic MCP server is a remote, cloud-hosted server that enables AI assistants like ChatGPT and Claude to convert articles, documents, websites, and videos into high-quality AI-generated audio. It provides multi-format support for text and binary files, natural-sounding text-to-audio conversion using AI, and specialized processing for SSML, markup, markdown, and various media formats through three core tools: listentic_supported_mimetypes, listentic_add_content_text, and listentic_add_content_binary.