MCP server for MarkItUp's AI image-annotation pipeline. Generate polished marketing-visual variations of any screenshot, regenerate, AI outpaint, and remove backgrounds —
powered by Claude analysis + Gemini rendering.
A streamlined MCP server for XMP metadata embedding with beautiful formatting and smart filename indicators, enabling metadata embedding, reading, validation, and report generation for lifestyle, product, and orbit schemas.
Give your coding agent marketing superpowers. Layers Marketing MCP researches your market, renders videos with AI personas, posts to TikTok, runs paid campaigns, and measures what worked; 27 tools, each priced before it runs.
MCP server for FreezeText — OCR anything on your Mac screen from Claude, Cursor, or any MCP client. Freeze the screen and extract text via Apple Vision (videos, popups, protected PDFs), OCR a region or a base64 image, and manage a searchable capture history. 12 tools. Bridge open-source (MIT), FreezeText app is free.
Enables text-only agents to process images by accepting image files, base64 data, or URLs, sending them to multimodal models, and returning structured text results via MCP.
An MCP server that recreates Microsoft Comic Chat (1996) to generate comic strips from conversation summaries, using original character art and deterministic layout, with no LLM calls.
Identifies LEGO parts, sets, and minifigures from local image files using the Brickognize API. It provides specialized tools for specific item recognition and integrates LEGO identification capabilities into MCP-enabled environments.
A local MCP server that provides image processing tools including resizing, cropping, format conversion, compression, rotation, flipping, thumbnailing, watermarking, effects, placeholder generation, and overlaying.
Enables AI image generation using Doubao Seedream models and video generation using Doubao Seedance models through Volcano Engine's API, supporting text-to-image, image-to-image, text-to-video, and task status queries.
Local stdio MCP service that calls GPT Image API to generate images from text, edit images with references or masks, and save images locally while returning absolute paths and file URIs.
Enables AI agents to perform deterministic, non-generative image transformations—straightening, cropping, masking, layering, color adjustments, and encoding—as reproducible recipes while preserving immutable originals.
Enables engraving LilyPond sources into cropped, self-contained EPS, PDF, SVG, and PNG assets for placement in page-layout software, with support for files or inline code and helpful diagnostics on failure.
Enables image generation and editing, video generation and extension, status polling, and download through the Kenari.id media APIs, saving outputs as files on disk.