Enables AI agents to perform detailed local image analysis, including pixel-level queries, cropping, grid coordinates, comparisons, and zero-shot classification, all without cloud APIs, keys, or payments.
Provides offline, privacy-preserving image recognition, OCR, and scene description for AI assistants via the Model Context Protocol, with Vulkan-accelerated local processing.
Provides local, offline transcription, keyframe extraction, OCR, and pre-publish review of audio, video, and image files, enabling AI agents to see and hear media without cloud or API keys.