inspect_pdf
Classify each PDF page as text, scanned, mixed, or blank using the text layer and image geometry, to identify which pages require OCR.
Instructions
Classify each page of a PDF as text / scanned / mixed / blank.
Uses the PDF text layer and image geometry only — no OCR is run. Use this first to decide whether (and which pages need) OCR.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| path | Yes |