Enables local OCR transcription of images using tesseract.js, with optional low-token AI-generated descriptions, folder batch processing, and compatibility with MCP clients like ChatGPT, Claude, opencode, and Cursor.
Enables AI assistants to manage Cloudinary media assets, including uploading, searching, transforming, and organizing images and videos through natural language.
Enables an AI assistant to batch-drive ArmorPaint headlessly via its native CLI and scripting engine, re-exporting projects at different presets and building procedural materials without opening the GUI.
Enables AI-driven 3D modeling and scene creation in Blender, allowing LLMs to create, modify, and manipulate objects, materials, and assets through natural language.
A customizable Model Context Protocol server implementation that enables AI models to interact with external tools including weather queries, Google search, and camera control functionality.
Enables detection of AI-generated content in images, videos, audio, and text via the AI or Not API. Supports media analysis tools for deepfakes, synthetic voices, and AI-written text.
Enables interaction with Figma designs through the Figma API, allowing users to export images in multiple formats, extract style data and CSS, analyze design elements, and retrieve SVG code from Figma files. Supports batch operations and comprehensive design element analysis including images, vectors, and components.
A server that allows AI tools like Claude and Cursor to automate and control Photoshop through natural language commands, enabling tasks like editing PSDs, toggling layers, and generating marketing creatives.
MCP server that exposes ComfyUI image and 3D generation as a single tool with smart prompt classification. It enables AI assistants to generate assets directly by selecting and running the appropriate ComfyUI workflow.
Enables image and video processing through FFmpeg, including compression, format conversion, resizing, and batch processing operations for common media formats.
A local MCP server for generative image description, providing prose captions, OCR, and LoRA dataset caption sidecars via Florence-2, with deterministic decoding and an honesty contract.
Enables AI-powered generation of styled QR codes with 10 design presets, supporting single or batch creation with custom logos and formats (SVG/PNG) directly from AI tools.
An MCP server for generating images and checking generation status through ComfyUI's API. It allows interaction with local ComfyUI instances by providing workflow-based image generation and status checking tools.
Provides tools to fetch IIIF manifests and retrieve specific image regions or scaled images for analysis. This server enables detailed interaction with International Image Interoperability Framework resources, supporting tasks like image description and transcription.