Skip to main content
Glama

glm-vision-mcp

An MCP server that wraps the Zhipu GLM-4.6V-Flash (free vision model), exposing an analyze_image tool to any MCP client, with support for single/multi-image analysis, OCR, and multi-image comparison.

Features

Capability

Description

Image analysis

Local paths / http(s) URLs / base64 data URIs are all supported, automatically converted to data URIs

Multi-image comparison

Pass multiple images in a single call and compare them according to the prompt

Rate-limit resilience

429 / 1302 / 1305 exponential backoff retry → multi-key rotation → fallback to backup model glm-4.1v-thinking-flash

Configuration self-check

The check_config tool verifies Key/model/endpoint without leaking the Key

Related MCP server: vision-mcp

Requirements

  • Python >= 3.10

  • Zhipu Open Platform API Key (https://open.bigmodel.cn/usercenter/apikeys), glm-4.6v-flash is free

  • To further reduce the chance of rate limiting, you can register multiple accounts and obtain multiple Keys, separated by commas in the configuration

Installation

cd glm-vision-mcp
python -m venv .venv
.venv\Scripts\pip install -r requirements.txt

Startup

# stdio 模式(MCP 客户端默认方式)
$env:ZHIPU_API_KEY = "你的Key"
.venv\Scripts\python server.py

# SSE 调试模式(无鉴权,仅限本机)
.venv\Scripts\python server.py --sse 8090

Client Configuration

Codex (~/.codex/config.toml)

[mcp_servers.glm-vision]
command = "C:\\绝对路径\\glm-vision-mcp\\.venv\\Scripts\\python.exe"
args = ["C:\\绝对路径\\glm-vision-mcp\\server.py"]

[mcp_servers.glm-vision.env]
ZHIPU_API_KEY = "你的Key"
# GLM_VISION_MODELS = "glm-4.6v-flash"
# GLM_API_BASE = "https://open.bigmodel.cn/api/paas/v4/chat/completions"

Claude Desktop (claude_desktop_config.json)

{
  "mcpServers": {
    "glm-vision": {
      "command": "C:\\绝对路径\\glm-vision-mcp\\.venv\\Scripts\\python.exe",
      "args": ["C:\\绝对路径\\glm-vision-mcp\\server.py"],
      "env": { "ZHIPU_API_KEY": "你的Key" }
    }
  }
}

Tool Interface

analyze_image(images, prompt, temperature, max_tokens, thinking)

Parameter

Type

Required

Description

images

string[]

Yes

Local path / http(s) URL / data URI

prompt

string

No

Analysis request, default "Please describe the content of this image in detail"

temperature

number

No

0.0~1.0, default 0.7

max_tokens

integer

No

Maximum output tokens, default 2048

thinking

boolean

No

Deep thinking mode, default false

Environment Variables

Variable

Required

Description

ZHIPU_API_KEY

Yes

Zhipu API Key, comma-separated for multi-key rotation

GLM_VISION_MODELS

No

Model priority, comma-separated, default glm-4.6v-flash,glm-4.1v-thinking-flash

GLM_API_BASE

No

Override API endpoint

Notes

  • Local images must be ≤ 10MB per image, supported formats: jpg/jpeg/png/webp/gif/bmp

  • Free models may be rate-limited during peak hours; an error is only raised when all Keys + all models are rate-limited. Wait 15~30 seconds and retry at off-peak times

  • Non-rate-limit errors such as 401 are not retried or downgraded; they are returned directly for easier troubleshooting

Verification

# 离线检查(不联网)
.venv\Scripts\python test_smoke.py

# 联网冒烟:MCP 握手 + analyze_image 真实调用
$env:ZHIPU_API_KEY = "你的Key"
.venv\Scripts\python test_smoke.py --live
Install Server
A
license - permissive license
A
quality
C
maintenance

Maintenance

Maintainers
Response time
Release cycle
Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    A
    quality
    A
    maintenance
    Multi-model vision understanding MCP server that provides unified image analysis for AI assistants without native vision, supporting models like GLM-4.6V, DeepSeek-OCR, Qwen3-VL-Flash, and more.
    1
    3,789
    100
    MIT
  • A
    license
    Not graded
    quality
    C
    maintenance
    MCP server for analyzing images using multiple vision LLM providers (OpenCode, OpenAI, Anthropic, Google, and custom OpenAI-compatible endpoints). Provides tools to analyze single or multiple images, list providers, and test vision capabilities.
    MIT

View all related MCP servers

Related MCP Connectors

  • MCP server for GLM chat completions using Zhipu AI models via AceDataCloud

  • OCR, transcription, file extraction, and image generation for AI agents via MCP.

  • MCP server for MiniMax H3 multimodal video generation

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/River831/glm-vision-mcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server