jev-eyesImage & Video ProcessingAI & Machine LearningLeddoEnganoAlicense-Not gradedqualityCmaintenanceEnables text-only agents to perceive images locally by converting screenshots into structured OCR text and spatial layout state, with optional zero-shot labeling and an ask tool for model decisions. Updated 3 days ago (2026-09-20 15:20 UTC)1MIT