Read document content (text / scan)
clio_document_readRead Clio documents without disk access: extract text from DOCX, PDF, TXT, EML, HTML with pagination, or return scanned pages as images for visual OCR.
Instructions
Returns the content of a Clio document directly in the response – no disk access needed. DOCX/PDF with a text layer/TXT/EML/HTML → text (paginate with offset/max_chars). Scanned PDFs without a text layer and images (JPG/PNG) → returns the pages as images that Claude reads (visual OCR); select pages with page_from/page_to (max 4 per call). mode: auto (default), text (text layer only), images (always page images – e.g. for stamps, signatures, tables). Accepts document_id or file_path.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| mode | No | ||
| offset | No | Character offset to continue from | |
| page_to | No | For page images: last page (max. 4 pages per call) | |
| file_path | No | Alternative to document_id – absolute path to a local file | |
| max_chars | No | Max. characters of text in the response (default 40000) | |
| page_from | No | For page images: first page (default 1) | |
| document_id | No | ||
| document_version_id | No |