A
licenseNot graded
qualityC
maintenanceProvides vision capabilities to text-only LLMs via MCP, enabling image understanding, Q&A, OCR, and image processing through cloud multimodal APIs.
MIT
No user-submitted related servers found.
Scored across 1 tool
There is only one tool, so there is no possibility of confusing it with another tool. Its purpose is clearly scoped to image analysis.
The single tool name follows a clear verb_noun pattern and there are no conflicting naming conventions to cause inconsistency.
One tool is at the low end, but it is justified for a narrow vision-analysis server. The built-in full/OCR modes keep it reasonably well-scoped rather than feeling empty.
For the stated purpose of providing image understanding to a non-vision model, the tool covers both general visual description and explicit OCR text/table extraction. There are no obvious missing core operations for this narrow domain.