llm-vision-mcp
Related Servers
Alternatives to llm-vision-mcp
No user-submitted related servers found.
Related Servers
- AlicenseAqualityAmaintenanceEnables text-only AI coding agents to analyze images and videos via vision-capable models (Gemini, Grok, OpenRouter), returning text descriptions for reasoning.213 npmMIT
- AlicenseAqualityBmaintenanceGives text-only coding agents the ability to 'see' images, videos, and screenshots by routing them to a vision model and returning structured text.813 npm1MIT
- AlicenseAqualityAmaintenanceEnables converting images (JPEG, PNG, GIF, WebP) into text descriptions using OpenAI-compatible vision models, with support for both local files and URLs.213 npmMIT
- FlicenseNot gradedqualityBmaintenanceEnables pure text LLMs to understand images by acting as a proxy to vision models via OpenAI-compatible APIs. Supports local files, URLs, and base64 inputs for image analysis.-
- AlicenseAqualityBmaintenanceGives vision-less LLMs the ability to recognize clipboard screenshots and images by proxying to an OpenAI-compatible vision model.112 npm2MIT
- FlicenseNot gradedqualityDmaintenanceEnables text-only language models to 'see' and describe images by calling multimodal APIs (OpenAI, Anthropic) for image analysis.-
TDQS
Scored across 1 tool
With only one tool, there is no possibility of confusion or overlap. The single tool has a clear and distinct purpose: analyzing images.
The tool name 'analyze_image' follows a clean verb_noun pattern. Since there is only one tool, naming consistency is perfect.
The server has exactly one tool, which feels slightly thin but is reasonable for a highly focused vision analysis server. The tool is comprehensive, handling many tasks through parameters, so the count is not inadequate.
The tool covers a wide range of vision tasks including describe, OCR, UI, layout, and QA, with multiple input sources and output options. There are no obvious gaps for the stated purpose of enabling text-only agents to analyze images.