Doubao Vision MCP Server
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
| DOUBAO_MODEL | No | 模型名称 | doubao-seed-2.0-pro |
| DOUBAO_API_KEY | Yes | 火山引擎 API Key | |
| DOUBAO_ENDPOINT | No | API 端点 | https://ark.cn-beijing.volces.com/api/v3/chat/completions |
| DOUBAO_MAX_TOKENS | No | 最大输出 token 数 | 1000 |
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Features and capabilities supported by this server
Protocol revision2025-11-25
| Capability | Details |
|---|---|
| tools | {
"listChanged": false
} |
| experimental | {} |
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
| recognize_imageA | 使用豆包视觉模型(doubao-seed-2.0-pro)识别图片内容。支持本地绝对路径和网络 URL。返回模型的文字描述,可用于物体识别、文字提取、场景理解等。不支持视频或非图片文件。 |
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
TDQS
Scored across 1 tool
Single tool with a clear, distinct purpose: image recognition. No ambiguity with other tools.
The tool name 'recognize_image' follows a clear verb_noun pattern. Though only one tool exists, the naming is well-structured and conventional.
One tool is on the thin side for a 'Vision' server, but it serves a focused purpose. The count is borderline acceptable.
The server name implies broader vision capabilities, yet only image recognition is provided. Missing tools for model info, image metadata, or other vision tasks create significant gaps.