Extract visible text from your screen using OCR. Returns text grouped by detected windows with bounding boxes. Use when you need code, terminal, chat, or document content without visual layout. Avoids sending data to the cloud.
AGPL 3.0
OCR for images and Korean ID documents
Scan any URL for AI agent readability — Vercel Spec, llmstxt.org, and agent-protocol manifests.