Enables any MCP-capable agent to perform vision tasks like describing images, answering questions, OCR, and comparing images using supported vision backends.
Provides image recognition capabilities to MCP clients by integrating with OpenAI-compatible vision models, supporting local images, URLs, multi-image comparison, and model listing.
Enables text-only LLMs to analyze images by bridging DeepSeek's web vision chat via MCP, supporting single/multiple image analysis, batch glob processing, and Windows screen capture for any MCP client.