"Compare all differences between two source code folders, including binary files" matching MCP connectors:
GET /v1/connectors – MCP directory API referenceMatching Connector Tools:
ScanSing reads printed sheet music. Give Claude a PDF or a photo of a score and it gets back MusicXML, the format MuseScore, Sibelius, Finale and Dorico open. It is the recogniser inside the ScanSing app for choir singers, offered as a connector. Once a score is recognised, Claude can answer questions about it without recognising it again: - what the key, time signature, tempo and length are, and each part's range, clefs and lyrics; - one voice on its own, for example just the alto line of a choir piece; - the score or a part moved to another key, with notes respelled for the new key signature; - a MIDI file, one track per part, to listen to or rehearse with; - a short-lived download link to open the result in a browser or notation app. It handles PDFs up to 40 pages and PNG, JPEG, HEIC, TIFF and WebP images, can recognise selected pages of a long PDF, and can split a songbook into one MusicXML per piece. A file can be given as a web link or a small image; in Claude Code, a file on your computer is sent through a one-time upload link. Recognition spends pages from your ScanSing plan, so it is the only tool that is not read-only. Everything else works on the stored result and is free. The first connection with a new email comes with 20 free pages; paid plans start at $9 a month for 300 pages. Scores and results are deleted about an hour after recognition and are never used for training.
Change lyrics in an existing song. Agents upload authorized MP3 audio, provide a 0.1–6 second phrase and replacement words, preview generated singing, and export MP3/WAV. No website signup required: buy credits via x402 with USDC, then use a scoped key. Caller supplies timing. Resumable CLI and Codex/Claude Code setup: https://lyricpatch.com/agents/integrations. Hear a real example: https://lyricpatch.com/agents/demo.
Turn recordings, videos and sheet music into playable, editable scores. Transcribe audio or video (files or links) to sheet music, read sheet-music photos and PDFs, convert MIDI, MusicXML and ABC, edit notes in conversation with previews and undo, and export PDF, MIDI, MusicXML, guitar TAB, jianpu and audio. Hosted remote server with OAuth sign-in; one instrument, voice or solo piano is free.
Transcribe Russian audio and video with timestamps and optional speaker labels from files or links.
Generate AI music via the Lacuna Music API from MCP clients like Claude Desktop & Code.
125+ browser tools for PDF, Image, Video, Audio, AI, Scanner. Files never leave your device.
Give ears to Claude/Openclaw/Hermes/Codex/Grok Bot. Voibe turns recordings into text your AI agent can work with. Ask your agent to transcribe a meeting, interview, call, lecture, podcast episode or voice memo. The raw transcript arrives in the chat with speaker labels, timestamps and a summary. Attach the file in the chat, or point at a file or folder in Claude Code, where a whole folder of recordings works in
AI transcription from URLs or files. 119 languages, diarization, SRT/VTT/text export.
Generate and edit images, video and audio on one AITOPIA account: 300+ AI models (image, video, speech, music, sound effects), plain-language multi-step editing with a price shown first, named tools (upscale, remove background, outpaint, reframe, motion control, voice change), ready video/audio tools (trim, captions, text overlays, narration, background music, merges), transcripts (SRT/VTT), voice cloning, projects and folders. Every paid tool takes dryRun to price a call before spending.
Transcribe audio and video files to text, SRT and VTT subtitles in 99 languages with Whisper.
Make AI-drawn video from chat: your LLM draws every frame in code; add narration, music and SFX.
Convert files too big or exotic for a sandbox: 140+ formats, batch, OCR, AI extraction, TTS/STT
Write an original song for one person: lyrics, two full versions and cover art in about a minute.
The Listenetic MCP server is a remote, cloud-hosted server that enables AI assistants like ChatGPT and Claude to convert articles, documents, websites, and videos into high-quality AI-generated audio. It provides multi-format support for text and binary files, natural-sounding text-to-audio conversion using AI, and specialized processing for SSML, markup, markdown, and various media formats through three core tools: listentic_supported_mimetypes, listentic_add_content_text, and listentic_add_content_binary.