Provides intelligent OCR and PDF processing capabilities that automatically detect whether PDFs contain digital text or scanned images and apply appropriate extraction methods. Supports text extraction, OCR processing, structure analysis, and batch operations.
Provides OCR capabilities to extract text from PDF documents using Tesseract, with support for multiple languages including English and Simplified Chinese.
Enables comprehensive PDF processing including text extraction, image extraction, and OCR capabilities for reading text within images across multiple languages.
Enables PDF reading through MCP with text extraction, Chinese OCR for scanned documents, page rendering, and keyword file search from any compatible MCP client.
Provides OCR capabilities to add searchable text layers to scanned PDFs, with tools to check OCR need, process single files or batch folders, and integrate with Zotero attachment storage.
Enables AI-powered extraction and analysis of PDF documents with 40+ specialized tools for text, tables, images, layout analysis, security assessment, and document intelligence. Supports both text-based and scanned PDFs with OCR capabilities.