Why this server?
Provides OCR from images via stdin/stdout or REST API, matching the search for image recognition and OCR.
Flicense-qualityCmaintenanceEnables AI assistants and applications to perform optical character recognition (OCR) from images via stdin/stdout streams or REST API endpoints.Why this server?
Local OCR server using PP-OCRv6 for fast text extraction and VL for document structure analysis, fitting OCR needs.
Alicense-qualityBmaintenanceA local OCR MCP server that extracts text from images using PP-OCRv6 for fast text extraction and VL-1.6 for document structure analysis, with automatic model routing and GPU detection.2MITWhy this server?
Local OCR (RapidOCR) plus optional Qwen-VL for image summarization, supporting file paths and clipboard images.
Flicense-qualityCmaintenanceLocal OCR (RapidOCR) plus optional Qwen-VL for image summarization, supporting both file paths and clipboard images.Why this server?
OCR images or PDFs using Mistral OCR API, suited for image text extraction.
Alicense-qualityDmaintenanceOCR images or pdfs, locally or by URLs by using Mistral OCR API (paid)40MITWhy this server?
High-accuracy OCR via Google Gemini API, extracting text from images and CAPTCHA processing.
Flicense-qualityDmaintenanceProvides OCR services powered by Google's Gemini API to extract text from images via file paths or base64 strings. It enables high-accuracy text recognition and CAPTCHA processing through simple MCP tools.7Why this server?
Enables OCR on images and PDFs, including full-page and region OCR, fitting the query.
Flicense-qualityDmaintenanceEnables OCR on images and PDFs, including full-page OCR, region OCR by description or bounding box, and caching with summary capabilities.Why this server?
Modular OCR with pluggable backends (Apple Vision, PaddleOCR) for text, tables, formulas, and charts.
Flicense-qualityBmaintenanceModular OCR MCP server with pluggable backends (Apple Vision, PaddleOCR, PaddleOCR-VL) for extracting text, tables, formulas, and charts from images.Why this server?
Multiple text-type recognition, handwriting recognition, and high-precision parsing of complex documents.

PaddleOCR MCP Serverofficial
Alicense-qualityAmaintenanceMultiple text-type recognition, handwriting recognition, and high-precision parsing of complex documents.87,476Apache 2.0Why this server?
Intelligent OCR and PDF processing, automatically detecting digital vs scanned text and applying extraction.
Alicense-qualityDmaintenanceProvides intelligent OCR and PDF processing capabilities that automatically detect whether PDFs contain digital text or scanned images and apply appropriate extraction methods. Supports text extraction, OCR processing, structure analysis, and batch operations.MIT