ocr_layout_vision
Extract text from images with layout analysis, returning each block with its bounding box coordinates.
Instructions
Extract text with layout analysis. Returns blocks with bounding boxes.
Backend: vision. Apple Vision OCR — fast on-device GPU/ANE inference (macOS 10.15+). Best for CJK + major European languages. Zero install on macOS.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| lang | No | ||
| mode | No | base | |
| path | Yes |