Modular OCR MCP server supporting Apple Vision, PaddleOCR, and PaddleOCR-VL backends. Enables text, layout, table, formula, and chart extraction from images via natural language.
MCP server for raster image manipulation using Pillow and OpenCV. Enables image operations such as resizing, cropping, rotating, format conversion, filtering, and auto-enhancement via MCP tools.
Converts PDFs to Markdown for AI optimization, with tools to detect PDF type (text-based, scanned, mixed, image-based) and support page ranges, raw/compact/page-break options.