deepseek-visionAI & Machine LearningImage & Video ProcessingSunlianwangAlicense-qualityBmaintenanceMCP server that gives text-only LLMs vision capabilities by using a free multimodal model to perceive images, audio, and video, returning text for the main model to reason with. Last updated 2026-08-05241MIT