Skip to main content
Glama

image_to_image

Transform an input image based on a text prompt to generate a new image, with adjustable size and quality.

Instructions

基于参考图片和文本描述生成新图片

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
sizeNo输出图片尺寸1024x1024
promptYes变换描述文本
qualityNo图片质量medium
save_pathYes保存的本地文件路径
input_imageYes输入图片的本地文件路径
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden of disclosing behavioral traits. It mentions generation but does not specify that files are created at the 'save_path', potential side effects, or required permissions. The file I/O behavior is implied but not explicitly stated.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is very short (one sentence) and to the point. It is efficient but may be overly brief for a tool with five parameters. No wasted words, but could benefit from additional context.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The tool generates an image and saves it to a file, but the description does not explain the output behavior or return value. Without an output schema, the agent must infer what the tool returns. This omission reduces completeness.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, so the schema already documents each parameter. The description adds overall tool purpose but does not elaborate on parameter meaning beyond what the schema provides. Baseline score of 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's purpose: generating new images based on a reference image and text description. It implies distinction from the sibling 'text_to_image' by mentioning a reference image input, which is unique to this tool.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description implies usage when a reference image and text transform are needed, but offers no explicit guidance on when not to use it or alternatives beyond the sibling name. No exclusion criteria or context signals are provided.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Install Server

Other Tools

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/IronManCantFix/image-mcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server