vision-augmentImage & Video ProcessingAI & Machine LearningCaoMeiYouRenAlicenseAqualityAmaintenanceEnables non-vision LLMs to understand images, extract text via OCR, and parse documents through a unified MCP interface, with local-first processing and optional OpenAI-compatible channels. Updated 4 hours ago (2026-08-11 15:31 UTC)3MIT