Skip to main content
Glama
604,258 tools. Updated 2026-09-23 18:25

"A tool for processing complex PDF documents with tables, charts, OCR, and images" matching MCP tools:

  • Extract figures, tables, and equations from PDF documents using layout detection. Returns base64-encoded images with metadata for academic papers or any PDF URL.
    Apache 2.0
  • Extract charts, tables, and diagrams from documents as inline images for AI analysis. Optionally run OCR for text extraction.
    MIT
  • Parse documents into structured formats (markdown, JSON, text, HTML). Extract text, tables, charts, formulas, and code blocks from PDF, DOCX, PPTX, HTML, MD, XLSX, images.
    MIT
  • Create professional PDF reports from structured JSON input with sections, tables, charts, and images. Returns the PDF as base64-encoded data.
    MIT
  • Extract OCR text from screenshots, documents, tables, terminal output, or code images. Convert visual content into machine-readable text for analysis.
    MIT
  • Converts documents between 100+ formats (Office, PDF, images, HTML) while preserving Carbone template tags. Supports PDF options like watermarks, passwords, or page ranges.
    Apache 2.0

Matching MCP Servers

Matching MCP Connectors

  • Extract readable text from images, screenshots, photos, and scanned PDFs via OCR. Returns reading-order text with confidence and optional bounding boxes for receipts, invoices, and forms.
    MIT
  • Translate whole files into another language, preserving layout, tables, images, and formulas. Works with documents, subtitles, images, audio, and video.
    Apache 2.0
  • Convert Markdown documents into styled PDF files. Choose a preset theme or custom JSON settings, add a title, and get a formatted PDF with support for headings, tables, code, and local images.
    MIT
  • Convert or process files with GuruPDF: transform between 100+ formats (PDF, Word, Excel, images, ebooks) or apply PDF tools (compress, merge, split, protect, OCR). Saves result next to input.
    MIT
  • Add documents to a collection by providing a URL for download, processing them for text extraction, and indexing them for semantic search.
    MIT
  • List documents in a Needle collection to check processing status, inventory available files, and verify document availability before searching.
    MIT
  • Extract text from images using offline OCR. Reads text from screenshots, documents, and photos via local file paths or base64 data.
    MIT
  • Extract figures, tables, and equations from PDF documents using layout detection. Returns base64-encoded images of detected elements with metadata from academic papers or any PDF URL.
    Apache 2.0
  • Verify identity documents by submitting front and optional back images. Get structured OCR data and authenticity checks for fraud prevention.
    MIT
  • Convert scanned PDFs or images into searchable PDFs by running OCR and adding an invisible text layer, enabling full-text search in documents that were previously unsearchable.
    MIT
  • Convert scanned PDFs into searchable, editable documents using OCR. Choose quality, language, and skip OCR when text is already digital.
    MIT