glm-vision-mcp
Related Servers
Alternatives to glm-vision-mcp
No user-submitted related servers found.
Related Servers
- AlicenseNot gradedqualityCmaintenanceEnables analysis of local images through Kimi (Moonshot AI) vision models via the MCP protocol, supporting features like OCR and long context understanding.17 npmMIT
- AlicenseAqualityCmaintenanceEnables AI assistants to recognize images and videos, and generate images (text-to-image, infographics, image-to-image, batch) through MCP tools.61Apache 2.0
- AlicenseNot gradedqualityBmaintenanceEnables to analyze images using GLM-4V vision model via the MCP protocol, supporting local file paths and Data URLs.22 npmMIT
- AlicenseNot gradedqualityBmaintenanceEnables text-only LLMs to analyze images by bridging DeepSeek's web vision chat via MCP, supporting single/multiple image analysis, batch glob processing, and Windows screen capture for any MCP client.1MIT
- AlicenseNot gradedqualityBmaintenanceProvides image understanding and OCR via GLM-4.6V-Flash, supporting URL, base64, and local file inputs. Enables AI assistants to analyze images and extract text from screenshots, documents, and more.29 npmMIT
- FlicenseNot gradedqualityCmaintenanceEnables image recognition using vision models via OpenAI-compatible APIs, supporting multiple platforms like OpenAI, DeepSeek, and Ollama.-
TDQS
Scored across 1 tool
With only one tool, there is no possibility of selecting between overlapping tools. The tool's purpose is clearly described as image description and OCR, so no ambiguity exists.
The single tool name 'describe_image' follows a clear verb_noun pattern, which is internally consistent. There are no other tools to create inconsistencies.
The server offers only one tool, which feels too thin for a vision-focused server. Even though the domain is narrow, typical vision servers provide separate operations for description, OCR, and question answering, making a single catch-all tool insufficient.
The tool covers image description and text extraction, but lacks common vision capabilities such as answering questions about an image or analyzing multiple images. The combined single tool creates a limitation where specific tasks cannot be addressed independently.