Skip to main content
Glama
459,989 tools. Updated 2026-08-17 14:47

"High-Quality PDF Extraction to Text with Tokenization and Accurate Processing of Complex Layouts" matching MCP tools:

  • Execute multi-step web scraping workflows with AI automation, navigating websites, interacting with forms, and extracting structured data for complex scenarios requiring user simulation.
    MIT
  • List available extraction templates to find a suitable one for your document type. Use team templates for processing; system templates require adding to your Suparse UI templates.
    MIT
  • Submit a document URL for automatic classification, structured field extraction, PII masking, and quality scoring. Returns a job ID for tracking processing status.
    MIT
  • Extracts text from SAM.gov opportunity attachments such as RFPs and SOWs using their download URLs. Returns plain text with extraction metadata, handling PDF and HTML formats.
    MIT
  • Submit a receipt image or PDF for asynchronous AI extraction of receipt data. Receive a job ID to poll for status, with optional webhook notification. Returns extracted data for further processing.
    MIT
  • Read raw source documents with automatic text extraction for PDF, DOCX, XLSX, PPTX, and plain text. Paginate by pages, sheet, or line offset.
    MIT

Matching MCP Servers

Matching MCP Connectors

  • Check the status of an asynchronous receipt extraction job. Provide the job ID to get the current status: pending, processing, completed, or failed.
    MIT
  • Add documents to a collection by providing a URL for download, processing them for text extraction, and indexing them for semantic search.
    MIT
  • Parse a PDF equipment manual and ingest it into the RCA knowledge base, enabling root cause analysis with indexed fault codes, part numbers, and parse quality.
    MIT
  • Retrieve Echo3s technical specifications—supported formats, AI processing details, output quality, credit system, and platform info.
    MIT
  • Specify a screen region with device pixel coordinates to extract text via OCR, returning JSON with text, confidence, and bounding box for targeted analysis.
    MIT
  • Extract data from receipt images and generate expense reports as PDF, DOCX, or ODT in one step, using AI extraction and customizable templates.
    MIT
  • Convert a local PDF file into an editable PowerPoint (PPTX) presentation. Provide the file path and optionally set quality, language, OCR, and merge preferences.
    MIT
  • Securely replaces sensitive plaintext values with vault tokens, enabling reversible detokenization while preserving data utility. Supports deterministic tokenization for consistent token mapping.
    GPL 3.0
  • Convert scanned PDFs into searchable, editable documents using OCR. Choose quality, language, and skip OCR when text is already digital.
    MIT
  • Convert a local PDF file into an Excel XLSX spreadsheet. Supports OCR, merging sheets, and extraction quality settings to suit your data needs.
    MIT
  • Replace specific text in a PDF with an image. Provide the PDF, image, text to find, and page range to place the image exactly where the text occurs.
    MIT