Ideogram V3 MCP Server
The Ideogram V3 MCP Server provides AI agents with direct access to Ideogram V3 image generation capabilities via RunAPI, enabling the following:
Generate images from text (
text_to_image): Create images from text prompts usingideogram-v3-text-to-imageorideogram-v3-charactermodels, with options for style, aspect ratio, output count, and rendering speed.Edit images (
edit_image): Modify existing images usingideogram-v3-editorideogram-v3-character-editmodels.Reframe images (
reframe_image): Resize or reframe existing images to different aspect ratios (1:1, 3:4, 9:16, 4:3, 16:9) usingideogram-v3-reframe.Remix images (
remix_image): Blend or remix source images (with optional style reference images) usingideogram-v3-remixorideogram-v3-character-remixmodels.Poll task status (
get_task): Fetch the current status and output URLs for any previously created task using its task ID.Check pricing (
check_pricing): Look up current costs for any Ideogram V3 model and endpoint — no API key required.Flexible task execution: All creation tools support a
waitflag to either poll synchronously until completion or return immediately with a task ID for async workflows.
The server exposes 7 model variants across 4 endpoints and is compatible with MCP hosts such as Claude Code, Cursor, Windsurf, VS Code, and others.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@Ideogram V3 MCP ServerGenerate an image of a serene beach at sunset."
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
Why This Package?
@runapi.ai/ideogram-v3-mcp is a focused Model Context Protocol server for the Ideogram V3 model line on RunAPI.
It gives MCP-compatible assistants direct access to 4 endpoints and 7 model variants without loading the full RunAPI catalog.
Use this per-model server when an agent should stay scoped to Ideogram V3. Use @runapi.ai/mcp when one assistant should discover every RunAPI model line.
Related MCP server: @runapi.ai/gpt-4o-image-mcp
Install
Add it to Claude Code:
claude mcp add ideogram-v3 -s user -- npx -y @runapi.ai/ideogram-v3-mcpUse project scope when the server should be shared with a repository:
claude mcp add ideogram-v3 -s project -- npx -y @runapi.ai/ideogram-v3-mcpCodex, Cursor, Windsurf, VS Code, Roo Code, and other MCP hosts can use the same stdio command:
{
"mcpServers": {
"ideogram-v3": {
"command": "npx",
"args": ["-y", "@runapi.ai/ideogram-v3-mcp"]
}
}
}check_pricing works before sign-in. For task creation and status polling, ask your assistant to call the login tool. It opens a browser login and saves credentials to ~/.config/runapi/config.json, the same file used by runapi login.
Headless and CI hosts can still set RUNAPI_API_KEY before starting the MCP host.
Ready-made examples are in examples/ for Claude, Cursor, Windsurf, VS Code, and Roo Code.
Tools
Tool | Auth | Purpose |
| Yes | Create an Ideogram V3 edit image task and optionally wait for a terminal status. Returns the task id, status, and output URLs. |
| Yes | Create an Ideogram V3 reframe image task and optionally wait for a terminal status. Returns the task id, status, and output URLs. |
| Yes | Create an Ideogram V3 remix image task and optionally wait for a terminal status. Returns the task id, status, and output URLs. |
| Yes | Create an Ideogram V3 text to image task and optionally wait for a terminal status. Returns the task id, status, and output URLs. |
| Yes | Fetch the current status and latest payload for an existing task. |
| No | Look up current pricing for a Ideogram V3 model and endpoint. |
Models
Ideogram V3 covers 7 model variants across 4 endpoints. Each tool accepts the models listed for it:
Tool | Models |
|
|
|
|
|
|
|
|
Model availability can change between releases. Use check_pricing or the Ideogram V3 model page for the current catalog view.
Agent Prompts
Ask your assistant in natural language; it can inspect pricing, create the task, and return the task id plus output URLs.
Create a task
Run an Ideogram V3 edit image task with RunAPI.The assistant can call check_pricing, then edit_image, and return the task id, status, and output URLs.
Submit without waiting
Create the task but don't wait for it to finish.The assistant calls the create tool with wait: false and returns the task id. Check on it later with get_task.
Check pricing before creating
Check current Ideogram V3 pricing, then create the task if it matches my request.The assistant calls check_pricing and can link to the Ideogram V3 model page for the canonical catalog entry.
Configuration
The server resolves auth in this order:
RUNAPI_API_KEYenvironment variable, useful for headless and CI hosts~/.config/runapi/config.json, created by the MCPlogintool orrunapi loginNo key, which still allows
check_pricing
The config file is normally managed by login. A pre-provisioned headless config can use:
{
"apiKey": "your_runapi_key"
}Do not commit real API keys.
Links
Resource | URL |
Ideogram V3 model page | |
npm package | |
GitHub repository | |
RunAPI MCP overview | |
RunAPI docs |
License
Licensed under the Apache License, Version 2.0.
Available Tools
7 toolscheck_pricingB
Look up RunAPI pricing for the ideogram-v3 model line.
| Name | Required | Description | Default |
|---|---|---|---|
| model | No | Model slug. Defaults to the line's primary model. | |
| action | No | Endpoint name. Defaults to the endpoint that offers the model. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description carries the full behavioral burden, and it discloses almost nothing: whether the call requires authentication (a 'login' sibling exists), whether lookups consume credits or are rate-limited, whether results are cached, or what the response looks like. 'Look up' weakly implies a read-only operation but that is inference, not disclosure.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
A single eleven-word sentence with the resource and scope front-loaded and zero filler. Nothing could be removed without losing meaning.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a two-parameter, all-optional read tool this is roughly adequate, but with no output schema the description is the only place return-value shape (currency, per-endpoint tiers, units) could be explained, and it is silent. The auth question raised by the 'login' sibling is also unaddressed.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100% and both parameters carry inline descriptions including their defaulting behavior, so the baseline of 3 applies. The description adds only the model-line context ('ideogram-v3') and no syntax or defaulting detail beyond the schema.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb ('look up') and resource ('pricing') scoped to the RunAPI ideogram-v3 model line, which clearly separates it from the action-oriented siblings (edit_image, text_to_image, etc.). It stops short of naming an alternative tool or restating the distinction explicitly, so it lands at clear-but-not-exemplary.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
There is no statement of when to call this versus the sibling endpoints, no prerequisites, and no exclusions. An agent must infer that this is a pre-flight cost check before invoking an image endpoint; nothing in the text confirms that.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
edit_imageB
Create a Ideogram V3 task on RunAPI (edit image). Returns a task id, status, and output URLs.
| Name | Required | Description | Default |
|---|---|---|---|
| seed | No | Declared type: integer. | |
| wait | No | Poll until the task reaches a terminal status. | |
| model | No | RunAPI model slug for this model line. | |
| style | No | Declared type: string. | |
| prompt | No | Declared type: string. | |
| mask_url | Yes | Declared type: string. | |
| timeout_ms | No | ||
| callback_url | No | Declared type: string. | |
| output_count | No | Declared type: integer. Known values: 1, 2, 3, 4. | |
| rendering_speed | No | Declared type: string. Known values: "turbo", "balanced", "quality". | |
| poll_interval_ms | No | ||
| source_image_url | Yes | Declared type: string. | |
| reference_image_urls | No | Declared type: array. | |
| enable_prompt_expansion | No | Declared type: boolean. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries the full behavioral burden. It does add genuinely useful context by disclosing that this is an asynchronous task-creation model returning a task id and status (implying get_task follow-up), which is not derivable from the schema. However, it omits permissions, whether an edit is destructive to the source, rate limits, and how the mask constrains the operation.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two short sentences with zero filler, and the core action is front-loaded before the return-value clause. Every clause earns its place.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
This is a 14-parameter tool with no output schema and no annotations, so the description should be doing far more work. It never explains the mask/source-image contract, the meaning of wait/poll_interval_ms, or the model/style options, leaving an agent under-informed for a fairly complex invocation.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 86% across 14 parameters, so the schema largely documents itself and the baseline is 3. The description adds no parameter-level meaning (no explanation of mask_url, source_image_url, prompt, style, or wait semantics) beyond what the schema already states.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
Names a specific verb and resource ('Create a Ideogram V3 task on RunAPI (edit image)') and states the return contract (task id, status, output URLs), so an agent knows this kicks off an image-editing job. It does not explicitly differentiate itself from the semantically nearby siblings remix_image and reframe_image, which keeps it short of a 5.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
There is no statement of when to choose this tool over remix_image, reframe_image, or text_to_image, and no prerequisites (e.g. that a source image and mask are mandatory) are called out. The parenthetical '(edit image)' implies the domain but gives no routing guidance.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
get_taskB
Fetch the current status and latest result payload for a ideogram-v3 task.
| Name | Required | Description | Default |
|---|---|---|---|
| action | Yes | Asynchronous endpoint the task was created on. | |
| task_id | Yes | Task id returned when the task was created. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description carries the full burden. It discloses what the tool returns (status plus latest result payload), which is useful, but says nothing about auth requirements, error behavior for unknown/expired task ids, or async polling semantics for a task that is not yet complete.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
A single front-loaded sentence with the operation and its return payload stated up front. No filler or redundancy; every clause earns its place.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
With no annotations and no output schema, the description must carry behavioral and return-value context. It names the return payload but omits polling/readiness and error-state behavior. Adequate for a simple two-parameter poll tool, but not complete.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, and both parameters (task_id, action) are documented in the schema, including the enum values for action. The description adds no additional parameter meaning, so the baseline 3 applies.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb (Fetch) and resource (current status and latest result payload) scoped to an 'ideogram-v3 task'. An agent can tell this is a status/result retrieval tool, distinct from the sibling creation endpoints. However, it never names or explicitly distinguishes itself from those siblings.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description implies it is used after creating a task, but gives no explicit when-to-use guidance, no mention of polling cadence, retry, or what to do while the task is incomplete. No alternatives or exclusions are named.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
loginA
Authenticate RunAPI by opening a browser PKCE login flow and saving the API key to ~/.config/runapi/config.json.
| Name | Required | Description | Default |
|---|---|---|---|
| force | No | Re-run browser login when the current credential comes from the local config file. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries the behavioral burden. It discloses the interactive browser flow and the file write side effect (config.json). However, it does not mention that it may overwrite existing credentials or that it could block waiting for user input, though these are implied.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is a single, well-structured sentence that front-loads the action ('Authenticate RunAPI') and provides necessary details without extraneous information.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a simple login tool with one optional parameter and no output schema, the description covers the core purpose and side effect. It lacks an explicit statement that this is a prerequisite for other tools, but that is implied.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100% (the only parameter 'force' has a description). The tool description adds no additional meaning about parameters beyond the schema, so the baseline of 3 applies.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool's purpose with a specific verb ('Authenticate'), target resource ('RunAPI'), method ('browser PKCE login flow'), and side effect (saving to config.json). It is distinct from sibling tools, none of which relate to authentication.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description implies usage (to authenticate RunAPI) but does not explicitly say when to run it (e.g., before other RunAPI tools) or when to use the 'force' parameter. Since there are no alternative auth tools among siblings, 'vs alternatives' is not applicable.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
reframe_imageB
Create a Ideogram V3 task on RunAPI (reframe image). Returns a task id, status, and output URLs.
| Name | Required | Description | Default |
|---|---|---|---|
| seed | No | Declared type: integer. | |
| wait | No | Poll until the task reaches a terminal status. | |
| model | No | RunAPI model slug for this model line. | |
| style | No | Declared type: string. Known values: "auto", "general", "realistic", "design". | |
| timeout_ms | No | ||
| aspect_ratio | No | Declared type: string. Known values: "1:1", "3:4", "9:16", "4:3", "16:9". | |
| callback_url | No | Declared type: string. | |
| output_count | No | Declared type: integer. Known values: 1, 2, 3, 4. | |
| rendering_speed | No | Declared type: string. Known values: "turbo", "balanced", "quality". | |
| poll_interval_ms | No | ||
| source_image_url | Yes | Declared type: string. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description carries the full behavioral burden. It does usefully disclose the asynchronous task pattern and the return shape (task id, status, output URLs), but it says nothing about auth/permission needs, polling defaults, or rate limits, which matter for an 11-parameter task-creation tool.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two short sentences with the action front-loaded and no filler. Efficient, though the parenthetical 'reframe image' slightly repeats the tool name rather than adding new information.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a 1-required / 11-total-parameter tool with no annotations and no output schema, the description helpfully covers the return values (task id, status, output URLs). But it omits the key operational context an agent needs: how this differs from sibling image tools and how the wait/polling parameters change behavior.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 82% (above the 80% baseline), so the schema already documents most parameters. The description adds no parameter meaning whatsoever, and several schema descriptions are low-value type restatements, but the high coverage justifies the baseline 3.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb (Create) and resource (Ideogram V3 task on RunAPI, specifically reframe image), so the agent knows this is an image-reframing operation rather than a generic task. However, it does not distinguish itself from close siblings like edit_image, remix_image, or text_to_image, leaving the agent to infer the boundary.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
There is no when-to-use guidance and no named alternative. With siblings such as edit_image, remix_image, and text_to_image present, the agent receives no signal about which one applies to a given request. This is the description's largest gap.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
remix_imageC
Create a Ideogram V3 task on RunAPI (remix image). Returns a task id, status, and output URLs.
| Name | Required | Description | Default |
|---|---|---|---|
| seed | No | Declared type: integer. | |
| wait | No | Poll until the task reaches a terminal status. | |
| model | No | RunAPI model slug for this model line. | |
| style | No | Declared type: string. Known values: "auto", "realistic", "fiction", "general", "design". | |
| prompt | No | Declared type: string. | |
| strength | No | Declared type: number. | |
| timeout_ms | No | ||
| aspect_ratio | No | Declared type: string. Known values: "1:1", "3:4", "9:16", "4:3", "16:9". | |
| callback_url | No | Declared type: string. | |
| output_count | No | Declared type: integer. Known values: 1, 2, 3, 4. | |
| negative_prompt | No | Declared type: string. | |
| rendering_speed | No | Declared type: string. Known values: "turbo", "balanced", "quality". | |
| poll_interval_ms | No | ||
| source_image_url | Yes | Declared type: string. | |
| reference_mask_urls | No | Declared type: array. | |
| reference_image_urls | No | Declared type: array. | |
| enable_prompt_expansion | No | Declared type: boolean. | |
| style_reference_image_urls | No | Declared type: array. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description carries the full behavioral burden. It discloses only the return shape (task id, status, output URLs); it does not explain that the call is asynchronous, how the 'wait'/'poll_interval_ms' polling behaves, what auth is required, or what happens with callback_url.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two short sentences with the operation stated first and the return values second; no filler. Slightly dense but front-loaded and efficient.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For an 18-parameter asynchronous task tool with no annotations and no output schema, the description usefully states the return values but omits the async/polling model and any usage context. It covers the minimal essentials while leaving notable behavioral gaps for such a complex tool.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 89%, so the schema already documents nearly all 18 parameters, which sets the baseline at 3. The description adds no parameter-level meaning beyond implying a source image is needed for the remix.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description names a specific verb (Create) plus resource (Ideogram V3 task, remix image) and clarifies it acts on a source image, which hints at the distinction from text_to_image. However, it does not explain how 'remix' differs from sibling edit_image or reframe_image, so an agent cannot fully route between the image-manipulation siblings from the description alone.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
There is no when-to-use or when-not-to-use guidance and no mention of the alternative image tools (edit_image, reframe_image, text_to_image). The agent is left to infer from the name which operation applies.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
text_to_imageB
Create a Ideogram V3 task on RunAPI (text to image). Returns a task id, status, and output URLs.
| Name | Required | Description | Default |
|---|---|---|---|
| seed | No | Declared type: integer. | |
| wait | No | Poll until the task reaches a terminal status. | |
| model | No | RunAPI model slug for this model line. | |
| style | No | Declared type: string. Known values: "auto", "realistic", "fiction", "general", "design". | |
| prompt | No | Declared type: string. | |
| timeout_ms | No | ||
| aspect_ratio | No | Declared type: string. Known values: "1:1", "3:4", "9:16", "4:3", "16:9". | |
| callback_url | No | Declared type: string. | |
| output_count | No | Declared type: integer. Known values: 1, 2, 3, 4. | |
| negative_prompt | No | Declared type: string. | |
| rendering_speed | No | Declared type: string. Known values: "turbo", "balanced", "quality". | |
| poll_interval_ms | No | ||
| reference_image_urls | No | Declared type: array. | |
| enable_prompt_expansion | No | Declared type: boolean. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries the full load. It does add real value by disclosing the async task model and the return shape (task id, status, output URLs), which no output schema supplies. However, it omits auth needs, cost/pricing implications, and how the wait/poll parameters interact with the returned task id.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two tight sentences with the action front-loaded and the return contract second. No filler, though the phrasing 'Create a Ideogram V3 task on RunAPI' is slightly awkward and the parenthetical repeats the tool name.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a 14-parameter generation tool with no annotations and no output schema, the description is thin. It covers the return contract but says nothing about required inputs despite prompt obviously being essential, cost checks, polling via get_task, or how the wait/timeout_ms/callback_url options change behavior.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 86%, so the schema already documents most of the 14 parameters, including enum-like known values for style, aspect_ratio, and rendering_speed. The description adds no parameter detail (e.g., model slug, output_count, callback_url) beyond that baseline.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb (Create) plus the resource (Ideogram V3 task via RunAPI) and disambiguates with the parenthetical '(text to image)'. That distinguishes it from image-editing siblings (edit_image, reframe_image, remix_image), though it never names them directly.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
There is no when-to-use guidance, no prerequisites, and no named alternative. The '(text to image)' tag hints at the domain but does not tell the agent when to pick this over remix_image or edit_image, nor that get_task/check_pricing may be needed alongside it.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
6 tool updates
v0.2.0- Changed
check_pricing2 fields changed- removed
Input schema / additionalPropertiesRemoved value: -false - removed
Input schema / properties / model / enumRemoved value: -[ - "ideogram-v3-character-edit", - "ideogram-v3-edit", - "ideogram-v3-reframe", - "ideogram-v3-character-remix", - "ideogram-v3-remix", - "ideogram-v3-character", - "ideogram-v3-text-to-image" -]
- Changed
edit_image18 fields changed- changed
Input schema / additionalPropertiesPrevious value: -falseNew value: +{} - added
Input schema / properties / callback_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / enable_prompt_expansion / descriptionAdded value: +"Declared type: boolean." - added
Input schema / properties / mask_url / descriptionAdded value: +"Declared type: string." - removed
Input schema / properties / model / enumRemoved value: -[ - "ideogram-v3-character-edit", - "ideogram-v3-edit" -] - added
Input schema / properties / output_count / descriptionAdded value: +"Declared type: integer. Known values: 1, 2, 3, 4." - removed
Input schema / properties / output_count / enumRemoved value: -[ - 1, - 2, - 3, - 4 -] - changed
Input schema / properties / output_count / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / poll_interval_ms / maximumAdded value: +9007199254740991 - added
Input schema / properties / prompt / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / reference_image_urls / descriptionAdded value: +"Declared type: array." - added
Input schema / properties / rendering_speed / descriptionAdded value: +"Declared type: string. Known values: \"turbo\", \"balanced\", \"quality\"." - removed
Input schema / properties / rendering_speed / enumRemoved value: -[ - "turbo", - "balanced", - "quality" -] - added
Input schema / properties / seed / descriptionAdded value: +"Declared type: integer." - changed
Input schema / properties / seed / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / source_image_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / style / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / timeout_ms / maximumAdded value: +9007199254740991
- Changed
get_task1 field changed- removed
Input schema / additionalPropertiesRemoved value: -false
- Changed
reframe_image17 fields changed- changed
Input schema / additionalPropertiesPrevious value: -falseNew value: +{} - added
Input schema / properties / aspect_ratio / descriptionAdded value: +"Declared type: string. Known values: \"1:1\", \"3:4\", \"9:16\", \"4:3\", \"16:9\"." - removed
Input schema / properties / aspect_ratio / enumRemoved value: -[ - "1:1", - "3:4", - "9:16", - "4:3", - "16:9" -] - added
Input schema / properties / callback_url / descriptionAdded value: +"Declared type: string." - removed
Input schema / properties / model / enumRemoved value: -[ - "ideogram-v3-reframe" -] - added
Input schema / properties / output_count / descriptionAdded value: +"Declared type: integer. Known values: 1, 2, 3, 4." - removed
Input schema / properties / output_count / enumRemoved value: -[ - 1, - 2, - 3, - 4 -] - changed
Input schema / properties / output_count / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / poll_interval_ms / maximumAdded value: +9007199254740991 - added
Input schema / properties / rendering_speed / descriptionAdded value: +"Declared type: string. Known values: \"turbo\", \"balanced\", \"quality\"." - removed
Input schema / properties / rendering_speed / enumRemoved value: -[ - "turbo", - "balanced", - "quality" -] - added
Input schema / properties / seed / descriptionAdded value: +"Declared type: integer." - changed
Input schema / properties / seed / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / source_image_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / style / descriptionAdded value: +"Declared type: string. Known values: \"auto\", \"general\", \"realistic\", \"design\"." - removed
Input schema / properties / style / enumRemoved value: -[ - "auto", - "general", - "realistic", - "design" -] - added
Input schema / properties / timeout_ms / maximumAdded value: +9007199254740991
- Changed
remix_image24 fields changed- changed
Input schema / additionalPropertiesPrevious value: -falseNew value: +{} - added
Input schema / properties / aspect_ratio / descriptionAdded value: +"Declared type: string. Known values: \"1:1\", \"3:4\", \"9:16\", \"4:3\", \"16:9\"." - removed
Input schema / properties / aspect_ratio / enumRemoved value: -[ - "1:1", - "3:4", - "9:16", - "4:3", - "16:9" -] - added
Input schema / properties / callback_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / enable_prompt_expansion / descriptionAdded value: +"Declared type: boolean." - removed
Input schema / properties / model / enumRemoved value: -[ - "ideogram-v3-character-remix", - "ideogram-v3-remix" -] - added
Input schema / properties / negative_prompt / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / output_count / descriptionAdded value: +"Declared type: integer. Known values: 1, 2, 3, 4." - removed
Input schema / properties / output_count / enumRemoved value: -[ - 1, - 2, - 3, - 4 -] - changed
Input schema / properties / output_count / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / poll_interval_ms / maximumAdded value: +9007199254740991 - added
Input schema / properties / prompt / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / reference_image_urls / descriptionAdded value: +"Declared type: array." - added
Input schema / properties / reference_mask_urls / descriptionAdded value: +"Declared type: array." - added
Input schema / properties / rendering_speed / descriptionAdded value: +"Declared type: string. Known values: \"turbo\", \"balanced\", \"quality\"." - removed
Input schema / properties / rendering_speed / enumRemoved value: -[ - "turbo", - "balanced", - "quality" -] - added
Input schema / properties / seed / descriptionAdded value: +"Declared type: integer." - changed
Input schema / properties / seed / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / source_image_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / strength / descriptionAdded value: +"Declared type: number." - added
Input schema / properties / style / descriptionAdded value: +"Declared type: string. Known values: \"auto\", \"realistic\", \"fiction\", \"general\", \"design\"." - removed
Input schema / properties / style / enumRemoved value: -[ - "auto", - "realistic", - "fiction", - "general", - "design" -] - added
Input schema / properties / style_reference_image_urls / descriptionAdded value: +"Declared type: array." - added
Input schema / properties / timeout_ms / maximumAdded value: +9007199254740991
- Changed
text_to_image21 fields changed- changed
Input schema / additionalPropertiesPrevious value: -falseNew value: +{} - added
Input schema / properties / aspect_ratio / descriptionAdded value: +"Declared type: string. Known values: \"1:1\", \"3:4\", \"9:16\", \"4:3\", \"16:9\"." - removed
Input schema / properties / aspect_ratio / enumRemoved value: -[ - "1:1", - "3:4", - "9:16", - "4:3", - "16:9" -] - added
Input schema / properties / callback_url / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / enable_prompt_expansion / descriptionAdded value: +"Declared type: boolean." - removed
Input schema / properties / model / enumRemoved value: -[ - "ideogram-v3-character", - "ideogram-v3-text-to-image" -] - added
Input schema / properties / negative_prompt / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / output_count / descriptionAdded value: +"Declared type: integer. Known values: 1, 2, 3, 4." - removed
Input schema / properties / output_count / enumRemoved value: -[ - 1, - 2, - 3, - 4 -] - changed
Input schema / properties / output_count / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / poll_interval_ms / maximumAdded value: +9007199254740991 - added
Input schema / properties / prompt / descriptionAdded value: +"Declared type: string." - added
Input schema / properties / reference_image_urls / descriptionAdded value: +"Declared type: array." - added
Input schema / properties / rendering_speed / descriptionAdded value: +"Declared type: string. Known values: \"turbo\", \"balanced\", \"quality\"." - removed
Input schema / properties / rendering_speed / enumRemoved value: -[ - "turbo", - "balanced", - "quality" -] - added
Input schema / properties / seed / descriptionAdded value: +"Declared type: integer." - changed
Input schema / properties / seed / typePrevious value: -"number"New value: +"integer" - added
Input schema / properties / style / descriptionAdded value: +"Declared type: string. Known values: \"auto\", \"realistic\", \"fiction\", \"general\", \"design\"." - removed
Input schema / properties / style / enumRemoved value: -[ - "auto", - "realistic", - "fiction", - "general", - "design" -] - added
Input schema / properties / timeout_ms / maximumAdded value: +9007199254740991 - added
Input schema / requiredAdded value: +[]
6 tool updates
v0.1.7- Changed
edit_image9 fields changed- added
Input schema / properties / callback_urlAdded value: +{ + "type": "string" +} - added
Input schema / properties / enable_prompt_expansionAdded value: +{ + "type": "boolean" +} - added
Input schema / properties / mask_urlAdded value: +{ + "type": "string" +} - added
Input schema / properties / promptAdded value: +{ + "type": "string" +} - added
Input schema / properties / reference_image_urlsAdded value: +{ + "items": {}, + "type": "array" +} - added
Input schema / properties / seedAdded value: +{ + "type": "number" +} - added
Input schema / properties / source_image_url / typeAdded value: +"string" - removed
Input schema / properties / style / enumRemoved value: -[ - "auto", - "realistic", - "fiction" -] - added
Input schema / requiredAdded value: +[ + "source_image_url", + "mask_url" +]
- Changed
get_task1 field changed- changed
Input schema / properties / action / descriptionPrevious value: -"Endpoint the task was created on."New value: +"Asynchronous endpoint the task was created on."
- Added
login - Changed
reframe_image4 fields changed- added
Input schema / properties / callback_urlAdded value: +{ + "type": "string" +} - added
Input schema / properties / seedAdded value: +{ + "type": "number" +} - added
Input schema / properties / source_image_url / typeAdded value: +"string" - added
Input schema / requiredAdded value: +[ + "source_image_url" +]
- Changed
remix_image12 fields changed- added
Input schema / properties / callback_urlAdded value: +{ + "type": "string" +} - added
Input schema / properties / enable_prompt_expansionAdded value: +{ + "type": "boolean" +} - added
Input schema / properties / negative_promptAdded value: +{ + "type": "string" +} - added
Input schema / properties / promptAdded value: +{ + "type": "string" +} - added
Input schema / properties / reference_image_urlsAdded value: +{ + "items": {}, + "type": "array" +} - added
Input schema / properties / reference_mask_urlsAdded value: +{ + "items": {}, + "type": "array" +} - added
Input schema / properties / seedAdded value: +{ + "type": "number" +} - added
Input schema / properties / source_image_url / typeAdded value: +"string" - added
Input schema / properties / strengthAdded value: +{ + "type": "number" +} - added
Input schema / properties / style_reference_image_urls / itemsAdded value: +{} - added
Input schema / properties / style_reference_image_urls / typeAdded value: +"array" - added
Input schema / requiredAdded value: +[ + "source_image_url" +]
- Changed
text_to_image6 fields changed- added
Input schema / properties / callback_urlAdded value: +{ + "type": "string" +} - added
Input schema / properties / enable_prompt_expansionAdded value: +{ + "type": "boolean" +} - added
Input schema / properties / negative_promptAdded value: +{ + "type": "string" +} - added
Input schema / properties / promptAdded value: +{ + "type": "string" +} - added
Input schema / properties / reference_image_urlsAdded value: +{ + "items": {}, + "type": "array" +} - added
Input schema / properties / seedAdded value: +{ + "type": "number" +}
6 tool updates
v0.1.0- First observed
check_pricing - First observed
edit_image - First observed
get_task - First observed
reframe_image - First observed
remix_image - First observed
text_to_image
TDQS
Scored across 7 tools
The four task-creation tools (text_to_image, edit_image, reframe_image, remix_image) all share near-identical descriptions and follow the same 'Create a Ideogram V3 task on RunAPI' phrasing, so an agent must infer the operational difference from the parenthetical label alone. They are genuinely distinct image operations, but the boilerplate wording makes boundaries less obvious than they should be.
All names use consistent snake_case with a clear verb_noun structure (edit_image, reframe_image, get_task, check_pricing, text_to_image). The lone 'login' is a natural single-verb auth tool and does not break the overall pattern.
Seven tools is well-scoped for an image-generation service, covering auth, four generation modes, status polling, and pricing without bloat or redundancy.
The surface covers the full generation lifecycle (auth, create for all four modes, poll via get_task, pricing). Minor gaps exist: no list_tasks or cancel_task to manage in-flight tasks, but core workflows are workable.
Maintenance
Related MCP Connectors
MCP server for Qwen Image 3 AI image generation
MCP server for Pixapi: check live credit pricing and balance, then generate images and video.
MCP server for Google Veo AI video generation
MCP server for Midjourney AI image generation and editing
Related MCP Servers
- AlicenseAqualityBmaintenanceOne MCP server for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one API key and one credit pool.16276 npm6MIT
- AlicenseBqualityAmaintenanceMCP server for GPT-4o Image model line, enabling text-to-image task creation, status polling, and pricing checks via RunAPI.4273 npmApache 2.0
- AlicenseAqualityAmaintenanceRunAPI MCP server for the Imagen 4 model line. Create tasks, poll their status, and check pricing through a single RunAPI API key.5254 npmApache 2.0
- AlicenseBqualityAmaintenanceMCP server for Veo 3.1 video generation models, enabling task creation (extend, text-to-video, upscale), status polling, and pricing checks via RunAPI.6260 npm1Apache 2.0