A
licenseA
qualityB
maintenanceMCP server that adds vision capabilities to text-only AI models by sending images (local files, URLs, clipboard, screenshots) to a vision model and returning text descriptions.
1
208 npm
MIT
No user-submitted related servers found.
Scored across 1 tool
With only one tool, there is no possibility of confusion or overlap. The tool's purpose is clearly defined in its description.
The single tool name 'describe_image' follows a clear verb_noun pattern, which is consistent and predictable even with only one tool.
A one-tool server feels thin, but the tool description covers a broad range of image understanding tasks. It is borderline, not excessive.
The tool covers many use cases (OCR, UI analysis, question answering, multiple image formats), but lacks separate operations like listing supported models or image preprocessing. Minor gaps, but core functionality is solid.