deepseek-visionAI & Machine LearningImage & Video ProcessingSunlianwangAlicense-Not gradedqualityBmaintenanceMCP server that gives text-only LLMs vision capabilities by using a free multimodal model to perceive images, audio, and video, returning text for the main model to reason with. Updated a month ago (2026-08-05 12:29 UTC)39 npm1MIT