Skip to main content
Glama

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
VI_MODELNoModel name (must support vision). Default: minimax-m3.minimax-m3
VI_API_KEYYesMiddleware API key (required).
VI_BASE_URLYesMiddleware base URL, e.g., https://api.xxx.com/v1 (required).
VI_MAX_TOKENSNoMaximum response tokens. Default: 2048.2048
VI_IMAGE_QUALITYNoJPEG quality (1-100). Default: 80.80
VI_MAX_IMAGE_SIZENoMax image width/height in pixels; larger images are scaled down. Default: 1280.1280
VI_MAX_IMAGE_BYTESNoMax compressed image size; errors if exceeded. Default: 10MB.10MB
VI_REQUEST_TIMEOUT_MSNoRequest timeout in milliseconds. Default: 60000.60000

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Features and capabilities supported by this server

Protocol revision2025-11-25

CapabilityDetails
tools
{}

Tools

Functions exposed to the LLM to take actions

NameDescription
analyze_imageA

分析本地图片(UI 自动化视觉辅助)。

当需要"查看屏幕/截图/界面"时使用:传入截图路径与问题,返回多模态模型的文字描述或结构化 JSON。

调用规范(必须遵守):

  • 调用前,先在回复中说明你正在分析哪张截图及其目的(如"让我分析一下当前界面")。

  • prompt 必须写明具体想知道的描述内容,例如 "描述界面布局并列出所有可见按钮"、"截图中有哪些错误提示"、"这个弹窗的标题和选项是什么"。

  • 禁止传空 prompt 或含糊 prompt(如 "看一下"、"描述" ),server 会把过短的 prompt 视为无效并自动补充标准描述要求。

  • 图片必须已保存为本地文件,传入绝对路径;本工具自行读取,无需传图片内容。

  • 需要结构化结果(如元素坐标)时设置 json_mode=true,要求模型返回 JSON。

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources

TDQS

A4.6/5.0

Scored across 1 tool

Disambiguation5/5

With only one tool, there is no possibility of confusion or overlap. The tool's purpose—analyzing images to answer questions or return structured data—is clear and unambiguous.

Naming Consistency5/5

The single tool name follows a clear verb_noun pattern (analyze_image) that is consistent and descriptive. There are no conflicting conventions to cause confusion.

Tool Count3/5

A single tool feels thin for a server branded as 'visual-intelligence,' but the tool itself is versatile enough to handle various image analysis requests. The count is right at the borderline where it could use additional specialized tools, but it is not wholly inappropriate.

Completeness4/5

The tool covers the core need of analyzing local images and returning either descriptive text or structured JSON, including support for UI-related queries. Minor gaps exist, such as requiring local file paths and lacking support for direct image URLs, but these are workaroundable.

Maintenance

ActivityMaintained
ResponsivenessNo issues