ocr_batch_paddleocr_vl
OCR multiple images at once and get consolidated results. Powered by PaddleOCR-VL, it supports 109 languages, tables, formulas, and charts.
Instructions
OCR multiple images at once. Returns consolidated results.
Backend: paddleocr_vl. PaddleOCR-VL — 0.9B vision-language model on Apple Silicon (M1+). Most accurate, 109 languages, supports tables/formulas/charts. Requires paddleocr-vl Swift CLI.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| lang | No | ||
| mode | No | base | |
| paths | Yes |