Request OCR scanning and structured data extraction for a package's document, label, envelope, or content. Receive extracted text, addresses, and dates.
Enables AI agents to search, deep-read, and build knowledge bases from Markdown, PDF, DOCX, and PPTX documents via MCP tools for retrieval, document navigation, and ingestion.
Extract text from images for document processing, receipt scanning, and image text extraction using OCR technology. Supports both URLs and base64 encoded images.
List available extraction templates to find a suitable one for your document type. Use team templates for processing; system templates require adding to your Suparse UI templates.
Submit a document URL for automatic classification, structured field extraction, PII masking, and quality scoring. Returns a job ID for tracking processing status.
Convert any document to markdown text with OCR for full text extraction, summarization, or translation. Provide a document ID, URL, or file data to get the complete markdown content.
Generate a browser upload link for adding a document to the workspace, returning a document_id for later extraction. Use when files can't be passed directly in chat.
Analyzes documents to automatically create JSON schemas for structured data extraction, enabling consistent field definitions across similar documents.
Update document properties in the content repository by providing the document identifier and new property values. Properties are updated without altering the document class.