Enables analysis of YouTube videos using the Gemini API to generate summaries and answer specific questions via direct URLs. It supports standard videos and shorts, allowing users to interact with video content without requiring manual downloads.
Enables an agent to inspect an audio folder and report what is inside each file — duration, sample rate, channels, codec and bitrate — and to group byte-identical files, identical decoded audio stored in different containers, and near-duplicate candidates, with suggested keeps and explicit differences, all read-only.
Provides Claude with detailed image inspection capabilities, including metadata extraction, histogram analysis, tonal and color analysis, sharpness detection, and more, supporting both standard and RAW formats.
Enables AI models to analyze audio files through numerical fingerprints, pitch tracking, and visual spectrograms without requiring direct audio playback. It provides tools for comparing audio iterations and detecting patterns using token-efficient analysis operations.
Enables Claude AI to control Ableton Live and Max for Live, allowing music production tasks like track management, MIDI editing, and pattern generation directly from conversation.
Enables generating beautiful syntax-highlighted code screenshots with professional themes directly from Claude. Supports file reading, line selection, git diff visualization, and batch processing across 20+ programming languages.
An MCP server that parses Douyin share links and performs intelligent content analysis using the Doubao video understanding model. It provides structured outputs including video summaries, categorized outlines, and step-by-step tutorial information.
Wallet-funded MCP client for six paid Utilia tools: Solana priority fees, transaction diagnosis and simulation, token-risk checks, PDF-to-Markdown, and audio normalization over x402. MIT-licensed and installable from npm.
Free Remotion video templates, creative guides and render diagnostics. No API key for local tools; optional paid hosted Studio workflows. Start with get_capabilities.
Enables Claude Code to use Alibaba Cloud Token Plan models via a local proxy that maps Anthropic-style model names to Alibaba IDs and provides missing endpoints like model discovery.
Enables AI assistants to locally process images with tools for cropping, zooming, enhancement, edge detection, segmentation, and text region extraction, all without external API keys. It uses PIL, OpenCV, and scikit-image for robust image analysis.
Enables generation of QR codes from text or URLs in multiple formats (DataURL, SVG, terminal display) with customizable options like error correction, colors, and size. Supports batch processing of multiple QR codes and integrates seamlessly with MCP-compatible clients.
Enables AI assistants and coding agents to automate Pictory.ai video creation, including 1080p rendering, transcription into viral short clips, AI avatars, stock asset search, project/template management, and publishing to Vimeo and S3.
Connects any MCP-compatible AI assistant to Audacity, providing 132 tools for real-time audio editing, cleanup, mastering, and transcription — all running locally without cloud dependencies.
A self-hosted MCP server that connects AI clients to local media-generation backends like ComfyUI and Blender, and provides built-in utilities for video/audio analysis, editing, subtitles, speech, and scene detection via a persistent file cache.