Skip to main content
Glama

Why This Package?

@runapi.ai/wan-mcp is a focused Model Context Protocol server for the Wan model line on RunAPI. It gives MCP-compatible assistants direct access to 6 endpoints and 18 model variants without loading the full RunAPI catalog.

Use this per-model server when an agent should stay scoped to Wan. Use @runapi.ai/mcp when one assistant should discover every RunAPI model line.


Related MCP server: @runapi.ai/gpt-4o-image-mcp

Install

Add it to Claude Code:

claude mcp add wan -s user -- npx -y @runapi.ai/wan-mcp

Use project scope when the server should be shared with a repository:

claude mcp add wan -s project -- npx -y @runapi.ai/wan-mcp

Codex, Cursor, Windsurf, VS Code, Roo Code, and other MCP hosts can use the same stdio command:

{
  "mcpServers": {
    "wan": {
      "command": "npx",
      "args": ["-y", "@runapi.ai/wan-mcp"]
    }
  }
}

check_pricing works before sign-in. For task creation and status polling, ask your assistant to call the login tool. It opens a browser login and saves credentials to ~/.config/runapi/config.json, the same file used by runapi login. Headless and CI hosts can still set RUNAPI_API_KEY before starting the MCP host.

Ready-made examples are in examples/ for Claude, Cursor, Windsurf, VS Code, and Roo Code.


Tools

Tool

Auth

Purpose

animate

Yes

Create a Wan animate task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

edit_video

Yes

Create a Wan edit video task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

image_to_video

Yes

Create a Wan image to video task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

speech_to_video

Yes

Create a Wan speech to video task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

text_to_image

Yes

Create a Wan text to image task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

text_to_video

Yes

Create a Wan text to video task and optionally wait for a terminal status. Returns the task id, status, and output URLs.

get_task

Yes

Fetch the current status and latest payload for an existing task.

check_pricing

No

Look up current pricing for a Wan model and endpoint.


Models

Wan covers 18 model variants across 6 endpoints. Each tool accepts the models listed for it:

Tool

Models

animate

wan-2.2-animate-move, wan-2.2-animate-replace

edit_video

wan-2.6-edit-video, wan-2.6-flash-edit-video, wan-2.7-edit-video

image_to_video

wan-2.2-a14b-image-to-video-turbo, wan-2.5-image-to-video, wan-2.6-flash-image-to-video, wan-2.6-image-to-video, wan-2.7-image-to-video

speech_to_video

wan-2.2-a14b-speech-to-video-turbo

text_to_image

wan-2.7-image, wan-2.7-image-pro

text_to_video

wan-2.2-a14b-text-to-video-turbo, wan-2.5-text-to-video, wan-2.6-text-to-video, wan-2.7-r2v, wan-2.7-text-to-video

Model availability can change between releases. Use check_pricing or the Wan model page for the current catalog view.


Agent Prompts

Ask your assistant in natural language; it can inspect pricing, create the task, and return the task id plus output URLs.

Create a task

Run a Wan animate task with RunAPI.

The assistant can call check_pricing, then animate, and return the task id, status, and output URLs.

Submit without waiting

Create the task but don't wait for it to finish.

The assistant calls the create tool with wait: false and returns the task id. Check on it later with get_task.

Check pricing before creating

Check current Wan pricing, then create the task if it matches my request.

The assistant calls check_pricing and can link to the Wan model page for the canonical catalog entry.


Configuration

The server resolves auth in this order:

  1. RUNAPI_API_KEY environment variable, useful for headless and CI hosts

  2. ~/.config/runapi/config.json, created by the MCP login tool or runapi login

  3. No key, which still allows check_pricing

The config file is normally managed by login. A pre-provisioned headless config can use:

{
  "apiKey": "your_runapi_key"
}

Do not commit real API keys.


Resource

URL

Wan model page

https://runapi.ai/models/wan

npm package

@runapi.ai/wan-mcp

GitHub repository

runapi-ai/wan-mcp

RunAPI MCP overview

runapi.ai/mcp

RunAPI docs

runapi.ai/docs


License

Licensed under the Apache License, Version 2.0.

Available Tools

9 tools
animateC

Create a Wan task on RunAPI (animate). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
waitNoPoll until the task reaches a terminal status.
modelNoRunAPI model slug for this model line.
timeout_msNo
callback_urlNoDeclared type: string.
poll_interval_msNo
source_image_urlYesDeclared type: string.
output_resolutionNoDeclared type: string. Known values: "480p", "580p", "720p".
reference_video_urlYesDeclared type: string.
enable_safety_checkerNoDeclared type: boolean.

TDQS

C2.8/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden, and it does disclose the async task pattern and return shape (task id, status, output URLs). However it never mentions the long-running/polling nature implied by the 'wait' and 'poll_interval_ms' parameters, nor cost, auth, or rate limits.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two tight sentences, front-loaded with the action and followed by the return shape; nothing is padded. It loses a point only because the terse phrasing leans on unexplained provider jargon rather than spending a few words on what the tool produces.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 9-parameter generative media tool with no annotations and no output schema, the description is far too thin. It explains the return fields but not the async lifecycle, the roles of the image/video inputs, or how polling/timeout interact, leaving an agent under-equipped to call it correctly.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 78%, just under the high-coverage threshold, and several parameters (source_image_url, reference_video_url, callback_url, enable_safety_checker) carry only 'Declared type' placeholders. The description adds no parameter meaning beyond the schema, so a baseline-adjacent 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose3/5

Does the description clearly state what the tool does and how it differs from similar tools?

It states a verb and resource ('Create a Wan task on RunAPI (animate)'), so the general action is clear. But it does not differentiate from the close siblings image_to_video and text_to_video, and 'animate' plus 'Wan task' is provider jargon rather than a description of what the animation actually does with a source image and reference video.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

There is no when-to-use guidance and no mention of alternatives, despite four visually similar siblings (image_to_video, text_to_video, edit_video, speech_to_video). An agent cannot infer from the text alone why it should pick 'animate' over image_to_video.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

check_pricingB

Look up RunAPI pricing for the wan model line.

ParametersJSON Schema
NameRequiredDescriptionDefault
modelNoModel slug. Defaults to the line's primary model.
actionNoEndpoint name. Defaults to the endpoint that offers the model.

TDQS

B3.1/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full disclosure burden. 'Look up' implies a read-only operation, but there is no mention of auth requirements, whether calls are billable/rate-limited, or the format of the pricing data returned — all relevant for a pricing endpoint.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

A single front-loaded sentence with zero filler; the resource and scope come first. Nothing repeats the name or wastes tokens.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

There is no output schema, so the description could usefully sketch the shape of returned pricing (units, currency, per-endpoint granularity) but does not. Scope is adequately stated, but return-value expectations are left unset for a tool whose whole purpose is retrieving data.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, so both 'model' and 'action' (including the action enum) are already documented in the schema. The description adds no syntax, defaulting, or relationship detail beyond what the schema states, so the baseline 3 applies.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description gives a specific verb ('look up') plus a bounded resource ('RunAPI pricing') and scope ('the wan model line'), so an agent knows exactly what the tool returns. It does not, however, distinguish itself from the sibling action tools or explain how the pricing scope relates to them, leaving that inference to the caller.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

There is no statement of when to call this versus alternatives, no mention of the defaulting behavior for model/action, and no guidance on what to do with the result. Only the verb 'look up' implies usage, which is weak guidance.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

edit_videoC

Create a Wan task on RunAPI (edit video). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
seedNoDeclared type: integer.
waitNoPoll until the task reaches a terminal status.
audioNoDeclared type: boolean.
modelNoRunAPI model slug for this model line.
promptNoDeclared type: string.
watermarkNoDeclared type: boolean.
timeout_msNo
multi_shotsNoControls whether the generated video uses multiple shots with transitions instead of one continuous shot. Declared type: boolean.
aspect_ratioNoDeclared type: string.
callback_urlNoDeclared type: string.
audio_settingNoDeclared type: string.
negative_promptNoDeclared type: string.
duration_secondsNoDeclared type: integer.
poll_interval_msNo
source_video_urlNoDeclared type: string.
output_resolutionNoDeclared type: string.
source_video_urlsNoDeclared type: array.
reference_image_urlNoDeclared type: string.
enable_safety_checkerNoDeclared type: boolean.
enable_prompt_expansionNoDeclared type: boolean.

TDQS

C2.6/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are present, so the description carries the full burden. It does disclose the async task pattern and the returned fields (task id, status, output URLs), which is useful, but it says nothing about required inputs (source_video_url?), whether wait=true blocks, cost/pricing implications, or failure behavior. For a 20-parameter generation tool with zero annotation coverage, that is a substantial gap.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two compact sentences with the operation stated first and the return contract second — no filler or repetition. It is appropriately sized, though the parenthetical "(edit video)" is doing less work than a real scope statement would.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

A 20-parameter, zero-required, no-annotation, no-output-schema tool needs more than two sentences. The description partially compensates for the missing output schema by naming the returned fields, but it leaves the agent unable to tell which of source_video_url / source_video_urls / reference_image_url are actually needed for an edit, nor how the wait/poll parameters interact.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 90%, so the schema already documents nearly all 20 parameters, and the description adds no parameter-level meaning beyond that. Baseline 3 is appropriate when the schema does the heavy lifting; the description contributes nothing extra here.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose3/5

Does the description clearly state what the tool does and how it differs from similar tools?

It gives a verb and resource ("Create a Wan task on RunAPI (edit video)") and names the return shape, but "edit video" is never qualified — it does not say whether it restyles an existing source video, concatenates clips, or performs some other transform. Siblings like text_to_video and image_to_video imply this line is for editing a source video, but the description never makes that distinction explicit.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

There is no when-to-use or when-not-to-use guidance at all. Nothing tells the agent how this differs from text_to_video, image_to_video, or animate, nor that get_task is the alternative for polling an existing task, even though this tool also exposes wait/poll_interval_ms.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

get_taskA

Fetch the current status and latest result payload for a wan task.

ParametersJSON Schema
NameRequiredDescriptionDefault
actionYesAsynchronous endpoint the task was created on.
task_idYesTask id returned when the task was created.

TDQS

A3.5/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full behavioral burden. It implies a read-only fetch and names the returned content at a high level, but omits auth requirements, polling behavior, status value semantics, rate limits, and error handling.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single front-loaded sentence with no wasted words. It states the action and returned payload immediately.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The tool is simple and the input schema is fully documented, but there is no output schema and no annotations. The description gives only a high-level view of the return payload and does not explain status values, polling expectations, or authentication needs, leaving gaps for a task-status tool.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, and the schema fully documents both parameters, including the action enum. The description adds no additional parameter meaning, so the baseline of 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description uses a specific verb ('Fetch') and resource ('current status and latest result payload for a wan task'), making it clearly distinct from the sibling creation and account tools. An agent can identify this as the task-status retrieval tool without opening the schema.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Usage is implied by 'current status' and 'latest result payload' for a task, suggesting it is used after an asynchronous task has been created. However, the description gives no explicit when-to-use guidance, prerequisites, or alternatives, leaving the agent to infer the polling context.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

image_to_videoB

Create a Wan task on RunAPI (image to video). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
seedNoDeclared type: integer.
waitNoPoll until the task reaches a terminal status.
audioNoDeclared type: boolean.
modelNoRunAPI model slug for this model line.
ratioNoDeclared type: string.
promptNoDeclared type: string.
watermarkNoDeclared type: boolean.
timeout_msNo
multi_shotsNoDeclared type: boolean.
accelerationNoDeclared type: string.
aspect_ratioNoDeclared type: string.
callback_urlNoDeclared type: string.
negative_promptNoDeclared type: string.
duration_secondsNoDeclared type: integer.
poll_interval_msNo
source_video_urlNoDeclared type: string.
driving_audio_urlNoDeclared type: string.
output_resolutionNoDeclared type: string. Known values: "480p", "720p", "1080p".
background_audio_urlNoDeclared type: string.
last_frame_image_urlNoDeclared type: string.
enable_safety_checkerNoDeclared type: boolean.
first_frame_image_urlNoDeclared type: string.
enable_prompt_expansionNoDeclared type: boolean.

TDQS

B3.1/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden, but it does disclose that this is an asynchronous task-creation call returning a task id, status, and output URLs. It says nothing about authentication, cost, rate limits, failure behavior, or what the 'wait' polling default implies; for a 23-parameter generation tool with zero annotation coverage this is only partially adequate.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two tight sentences with the verb and resource front-loaded and the return shape compactly stated; nothing is wasted. It is arguably under-specified for the tool's complexity, but that is a completeness concern rather than a verbosity one.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a task with 23 parameters, no required fields, no annotations, and no output schema, the description leaves major gaps: which image fields to supply, how model/ratio/duration interact, and what the polling behavior actually does. The return-value sentence is the only substantive addition, and it is not enough.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is reported at 91%, so the baseline is 3 even though the description adds no parameter-level meaning. In practice most schema entries are placeholders ('Declared type: integer'), so the description does nothing to explain key knobs like first_frame_image_url, last_frame_image_url, ratio, or wait.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb (create) and resource (a Wan image-to-video task on RunAPI), and the parenthetical 'image to video' distinguishes it from text_to_video, edit_video, and speech_to_video siblings. However, it never clarifies which input fields constitute the 'image' input, which matters given the tool has zero required parameters.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

No when-to-use guidance, no conditions, and no mention of alternatives such as text_to_video or animate, despite several closely related siblings. The agent must infer selection purely from the tool name.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

loginA

Authenticate RunAPI by opening a browser PKCE login flow and saving the API key to ~/.config/runapi/config.json.

ParametersJSON Schema
NameRequiredDescriptionDefault
forceNoRe-run browser login when the current credential comes from the local config file.

TDQS

A4/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the behavioral burden. It discloses the interactive browser flow and the file write side effect (config.json). However, it does not mention that it may overwrite existing credentials or that it could block waiting for user input, though these are implied.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single, well-structured sentence that front-loads the action ('Authenticate RunAPI') and provides necessary details without extraneous information.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a simple login tool with one optional parameter and no output schema, the description covers the core purpose and side effect. It lacks an explicit statement that this is a prerequisite for other tools, but that is implied.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100% (the only parameter 'force' has a description). The tool description adds no additional meaning about parameters beyond the schema, so the baseline of 3 applies.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's purpose with a specific verb ('Authenticate'), target resource ('RunAPI'), method ('browser PKCE login flow'), and side effect (saving to config.json). It is distinct from sibling tools, none of which relate to authentication.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description implies usage (to authenticate RunAPI) but does not explicitly say when to run it (e.g., before other RunAPI tools) or when to use the 'force' parameter. Since there are no alternative auth tools among siblings, 'vs alternatives' is not applicable.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

speech_to_videoC

Create a Wan task on RunAPI (speech to video). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
seedNoDeclared type: integer.
waitNoPoll until the task reaches a terminal status.
modelNoRunAPI model slug for this model line.
shiftNoDeclared type: number.
promptYesDeclared type: string.
num_framesNoDeclared type: integer.
timeout_msNo
callback_urlNoDeclared type: string.
guidance_scaleNoDeclared type: number.
negative_promptNoDeclared type: string.
poll_interval_msNo
source_audio_urlYesDeclared type: string.
source_image_urlYesDeclared type: string.
frames_per_secondNoDeclared type: integer.
output_resolutionNoDeclared type: string. Known values: "480p", "580p", "720p".
num_inference_stepsNoDeclared type: integer.
enable_safety_checkerNoDeclared type: boolean.

TDQS

C2.9/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full behavioral burden. It does disclose the return shape (task id, status, output URLs), which is useful, but says nothing about async/long-running generation, polling behavior implied by the 'wait' default, timeouts, cost, or auth requirements for the RunAPI call.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two short sentences, front-loaded with the action and immediately followed by the return shape. Nothing is padded or redundant.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 17-parameter media-generation tool with no annotations and no output schema, the description is too thin. It should cover the async task lifecycle, polling/timeout semantics, and the fact that generated video output is the payoff, none of which appear.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 88%, so the schema itself already documents the parameters and the baseline is 3. The description adds no parameter meaning at all – notably it doesn't mention that both a source image and a source audio URL are required.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource – 'Create a Wan task ... (speech to video)' – and the parenthetical modality distinguishes it from image_to_video and text_to_video. However it never says what inputs drive the generation (image + audio), leaving the agent to infer the distinction from the schema rather than the description.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

There is no when-to-use guidance and no reference to any alternative sibling tool. An agent cannot tell from the description why it would pick speech_to_video over animate or image_to_video.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

text_to_imageB

Create a Wan task on RunAPI (text to image). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
seedNoDeclared type: integer.
waitNoPoll until the task reaches a terminal status.
modelNoRunAPI model slug for this model line.
promptNoDeclared type: string.
bbox_listNoDeclared type: array.
watermarkNoDeclared type: boolean.
timeout_msNo
aspect_ratioNoDeclared type: string. Known values: "1:1", "16:9", "4:3", "21:9", "3:4", "9:16", "8:1", "1:8".
callback_urlNoDeclared type: string.
output_countNoDeclared type: integer.
color_paletteNoDeclared type: array.
thinking_modeNoDeclared type: boolean.
poll_interval_msNo
enable_sequentialNoDeclared type: boolean.
output_resolutionNoDeclared type: string. Known values: "1k", "2k", "4k".
source_image_urlsNoDeclared type: array.
enable_safety_checkerNoDeclared type: boolean.

TDQS

B3.1/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations exist, so the description carries the burden. It does disclose the return shape (task id, status, output URLs), which is valuable with no output schema, and implies an async task model. However, it never explains the polling behavior implied by 'wait', 'timeout_ms', and 'poll_interval_ms', nor cost, auth, or rate-limit considerations.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two short sentences with no filler and the action front-loaded. The jargon ('Wan task on RunAPI') is slightly opaque but does not waste words.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 17-parameter generative tool with zero annotations and no output schema, this is thin. Return values are mentioned, which helps, but there is nothing about async/wait semantics, valid model slugs, cost (despite a check_pricing sibling), or how to poll with get_task.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 88%, so the schema already documents the parameters, and the description adds no parameter-level meaning. Baseline 3 is appropriate when the schema does the heavy lifting.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a clear verb and resource: creating a text-to-image Wang task via RunAPI, and the parenthetical '(text to image)' distinguishes it from the sibling text_to_video. It stops short of naming alternatives explicitly, but the purpose is unambiguous.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

No when-to-use guidance, no exclusions, and no mention of the sibling tools an agent must choose between (text_to_video, edit_video, animate). Nothing tells the agent which model line or scenario this tool is appropriate for.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

text_to_videoB

Create a Wan task on RunAPI (text to video). Returns a task id, status, and output URLs.

ParametersJSON Schema
NameRequiredDescriptionDefault
seedNoDeclared type: integer.
waitNoPoll until the task reaches a terminal status.
modelNoRunAPI model slug for this model line.
ratioNoDeclared type: string.
promptNoDeclared type: string.
watermarkNoDeclared type: boolean.
timeout_msNo
multi_shotsNoDeclared type: boolean.
accelerationNoDeclared type: string.
aspect_ratioNoDeclared type: string.
callback_urlNoDeclared type: string.
negative_promptNoDeclared type: string.
duration_secondsNoDeclared type: integer.
poll_interval_msNo
output_resolutionNoDeclared type: string. Known values: "480p", "580p", "720p", "1080p".
reference_audio_urlNoDeclared type: string.
background_audio_urlNoDeclared type: string.
reference_image_urlsNoDeclared type: array.
reference_video_urlsNoDeclared type: array.
enable_safety_checkerNoDeclared type: boolean.
first_frame_image_urlNoDeclared type: string.
enable_prompt_expansionNoDeclared type: boolean.

TDQS

B3.4/5.0
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden. It does disclose the async task model ('Create a Wan task') and the return shape (task id, status, output URLs), which is genuinely useful. It omits cost implications, expected generation latency, auth requirements, and whether the task is billable or cancelable.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two sentences with zero padding, front-loaded with the action and platform, then the return values. Nothing is wasted and nothing needs reordering.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 22-parameter, zero-required tool with no annotations and no output schema, the description is thin. Naming the returned fields partially compensates for the missing output schema, but it says nothing about defaults, polling behavior tied to the 'wait' flag, or which parameters matter most, leaving an agent to guess at invocation.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 91%, so the schema nominally owns parameter documentation, which sets the baseline at 3. In practice those schema descriptions are largely filler ('Declared type: string'), and the tool description adds no parameter meaning at all — notably absent is any hint that prompt is the key input while the other 21 fields are optional knobs.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource: 'Create a Wan task on RunAPI (text to video)'. The parenthetical modality cleanly separates it from image_to_video and speech_to_video siblings, though those siblings are never named directly.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Usage is only implied: the '(text to video)' qualifier tells the agent to pick this when the input is a prompt rather than an image, audio clip, or existing video. There is no explicit when-to-use/when-not statement, no mention of the login prerequisite, and no comparison against edit_video or animate.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 8 tool updatesv0.2.0
    • Changedanimate10 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.2-animate-move",
        -  "wan-2.2-animate-replace"
        -]
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string. Known values: \"480p\", \"580p\", \"720p\"."
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "480p",
        -  "580p",
        -  "720p"
        -]
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / reference_video_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / source_image_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
    • Changedcheck_pricing2 fields changed
      • removedInput schema / additionalProperties
        Removed value: -false
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.2-animate-move",
        -  "wan-2.2-animate-replace",
        -  "wan-2.6-edit-video",
        -  "wan-2.6-flash-edit-video",
        -  "wan-2.7-edit-video",
        -  "wan-2.2-a14b-image-to-video-turbo",
        -  "wan-2.5-image-to-video",
        -  "wan-2.6-flash-image-to-video",
        -  "wan-2.6-image-to-video",
        -  "wan-2.7-image-to-video",
        -  "wan-2.2-a14b-speech-to-video-turbo",
        -  "wan-2.7-image",
        -  "wan-2.7-image-pro",
        -  "wan-2.2-a14b-text-to-video-turbo",
        -  "wan-2.5-text-to-video",
        -  "wan-2.6-text-to-video",
        -  "wan-2.7-r2v",
        -  "wan-2.7-text-to-video"
        -]
    • Changededit_video23 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / aspect_ratio / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / audio / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / audio_setting / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / duration_seconds / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / duration_seconds / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / enable_prompt_expansion / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.6-edit-video",
        -  "wan-2.6-flash-edit-video",
        -  "wan-2.7-edit-video"
        -]
      • addedInput schema / properties / multi_shots / description
        Added value: +"Controls whether the generated video uses multiple shots with transitions instead of one continuous shot. Declared type: boolean."
      • addedInput schema / properties / negative_prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / reference_image_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / seed / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / seed / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / source_video_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / source_video_urls / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / watermark / description
        Added value: +"Declared type: boolean."
      • addedInput schema / required
        Added value: +[]
    • Changedget_task1 field changed
      • removedInput schema / additionalProperties
        Removed value: -false
    • Changedimage_to_video27 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / acceleration / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / aspect_ratio / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / audio / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / background_audio_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / driving_audio_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / duration_seconds / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / duration_seconds / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / enable_prompt_expansion / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / first_frame_image_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / last_frame_image_url / description
        Added value: +"Declared type: string."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.2-a14b-image-to-video-turbo",
        -  "wan-2.5-image-to-video",
        -  "wan-2.6-flash-image-to-video",
        -  "wan-2.6-image-to-video",
        -  "wan-2.7-image-to-video"
        -]
      • addedInput schema / properties / multi_shots / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / negative_prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string. Known values: \"480p\", \"720p\", \"1080p\"."
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "480p",
        -  "720p",
        -  "1080p"
        -]
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / ratio / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / seed / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / seed / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / source_video_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / watermark / description
        Added value: +"Declared type: boolean."
      • addedInput schema / required
        Added value: +[]
    • Changedspeech_to_video22 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / frames_per_second / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / frames_per_second / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / guidance_scale / description
        Added value: +"Declared type: number."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.2-a14b-speech-to-video-turbo"
        -]
      • addedInput schema / properties / negative_prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / num_frames / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / num_frames / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / num_inference_steps / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / num_inference_steps / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string. Known values: \"480p\", \"580p\", \"720p\"."
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "480p",
        -  "580p",
        -  "720p"
        -]
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / seed / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / seed / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / shift / description
        Added value: +"Declared type: number."
      • addedInput schema / properties / source_audio_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / source_image_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
    • Changedtext_to_image22 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / aspect_ratio / description
        Added value: +"Declared type: string. Known values: \"1:1\", \"16:9\", \"4:3\", \"21:9\", \"3:4\", \"9:16\", \"8:1\", \"1:8\"."
      • removedInput schema / properties / aspect_ratio / enum
        Removed value: -[
        -  "1:1",
        -  "16:9",
        -  "4:3",
        -  "21:9",
        -  "3:4",
        -  "9:16",
        -  "8:1",
        -  "1:8"
        -]
      • addedInput schema / properties / bbox_list / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / color_palette / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / enable_sequential / description
        Added value: +"Declared type: boolean."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.7-image",
        -  "wan-2.7-image-pro"
        -]
      • addedInput schema / properties / output_count / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / output_count / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string. Known values: \"1k\", \"2k\", \"4k\"."
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "1k",
        -  "2k",
        -  "4k"
        -]
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / seed / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / seed / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / source_image_urls / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / thinking_mode / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / watermark / description
        Added value: +"Declared type: boolean."
      • addedInput schema / required
        Added value: +[]
    • Changedtext_to_video26 fields changed
      • changedInput schema / additionalProperties
        Previous value: -falseNew value: +{}
      • addedInput schema / properties / acceleration / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / aspect_ratio / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / background_audio_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / callback_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / duration_seconds / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / duration_seconds / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / enable_prompt_expansion / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / enable_safety_checker / description
        Added value: +"Declared type: boolean."
      • addedInput schema / properties / first_frame_image_url / description
        Added value: +"Declared type: string."
      • removedInput schema / properties / model / enum
        Removed value: -[
        -  "wan-2.2-a14b-text-to-video-turbo",
        -  "wan-2.5-text-to-video",
        -  "wan-2.6-text-to-video",
        -  "wan-2.7-r2v",
        -  "wan-2.7-text-to-video"
        -]
      • addedInput schema / properties / multi_shots
        Added value: +{
        +  "description": "Declared type: boolean.",
        +  "type": "boolean"
        +}
      • addedInput schema / properties / negative_prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / output_resolution / description
        Added value: +"Declared type: string. Known values: \"480p\", \"580p\", \"720p\", \"1080p\"."
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "480p",
        -  "580p",
        -  "720p",
        -  "1080p"
        -]
      • addedInput schema / properties / poll_interval_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / prompt / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / ratio / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / reference_audio_url / description
        Added value: +"Declared type: string."
      • addedInput schema / properties / reference_image_urls / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / reference_video_urls / description
        Added value: +"Declared type: array."
      • addedInput schema / properties / seed / description
        Added value: +"Declared type: integer."
      • changedInput schema / properties / seed / type
        Previous value: -"number"New value: +"integer"
      • addedInput schema / properties / timeout_ms / maximum
        Added value: +9007199254740991
      • addedInput schema / properties / watermark / description
        Added value: +"Declared type: boolean."
      • addedInput schema / required
        Added value: +[]
  2. 7 tool updatesv0.1.9
    • Changedanimate5 fields changed
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / reference_video_url / type
        Added value: +"string"
      • addedInput schema / properties / source_image_url / type
        Added value: +"string"
      • addedInput schema / required
        Added value: +[
        +  "source_image_url",
        +  "reference_video_url"
        +]
    • Changededit_video17 fields changed
      • removedInput schema / properties / aspect_ratio / enum
        Removed value: -[
        -  "16:9",
        -  "9:16",
        -  "1:1",
        -  "4:3",
        -  "3:4"
        -]
      • addedInput schema / properties / audio
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / audio_setting
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / duration_seconds
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / enable_prompt_expansion
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / multi_shots
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / negative_prompt
        Added value: +{
        +  "type": "string"
        +}
      • removedInput schema / properties / output_resolution / enum
        Removed value: -[
        -  "720p",
        -  "1080p"
        -]
      • addedInput schema / properties / prompt / type
        Added value: +"string"
      • addedInput schema / properties / reference_image_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / seed
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / source_video_url / type
        Added value: +"string"
      • addedInput schema / properties / source_video_urls / items
        Added value: +{}
      • addedInput schema / properties / source_video_urls / type
        Added value: +"array"
      • addedInput schema / properties / watermark
        Added value: +{
        +  "type": "boolean"
        +}
    • Changedget_task1 field changed
      • changedInput schema / properties / action / description
        Previous value: -"Endpoint the task was created on."New value: +"Asynchronous endpoint the task was created on."
    • Changedimage_to_video18 fields changed
      • addedInput schema / properties / acceleration
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / aspect_ratio
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / audio / type
        Added value: +"boolean"
      • addedInput schema / properties / background_audio_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / driving_audio_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / duration_seconds / type
        Added value: +"number"
      • addedInput schema / properties / enable_prompt_expansion
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / first_frame_image_url / type
        Added value: +"string"
      • addedInput schema / properties / last_frame_image_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / multi_shots
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / negative_prompt
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / prompt / type
        Added value: +"string"
      • addedInput schema / properties / ratio
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / seed
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / source_video_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / watermark
        Added value: +{
        +  "type": "boolean"
        +}
    • Changedspeech_to_video13 fields changed
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / frames_per_second
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / guidance_scale
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / negative_prompt
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / num_frames
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / num_inference_steps
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / prompt / type
        Added value: +"string"
      • addedInput schema / properties / seed
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / shift
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / source_audio_url / type
        Added value: +"string"
      • addedInput schema / properties / source_image_url / type
        Added value: +"string"
      • addedInput schema / required
        Added value: +[
        +  "source_image_url",
        +  "source_audio_url",
        +  "prompt"
        +]
    • Changedtext_to_image11 fields changed
      • addedInput schema / properties / bbox_list
        Added value: +{
        +  "items": {},
        +  "type": "array"
        +}
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / color_palette
        Added value: +{
        +  "items": {},
        +  "type": "array"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / enable_sequential
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / output_count
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / prompt
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / seed
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / source_image_urls
        Added value: +{
        +  "items": {},
        +  "type": "array"
        +}
      • addedInput schema / properties / thinking_mode
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / watermark
        Added value: +{
        +  "type": "boolean"
        +}
    • Changedtext_to_video17 fields changed
      • addedInput schema / properties / acceleration
        Added value: +{
        +  "type": "string"
        +}
      • removedInput schema / properties / aspect_ratio / enum
        Removed value: -[
        -  "16:9",
        -  "9:16",
        -  "1:1",
        -  "4:3",
        -  "3:4"
        -]
      • addedInput schema / properties / background_audio_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / callback_url
        Added value: +{
        +  "type": "string"
        +}
      • removedInput schema / properties / duration_seconds / maximum
        Removed value: -10
      • removedInput schema / properties / duration_seconds / minimum
        Removed value: -2
      • addedInput schema / properties / enable_prompt_expansion
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / enable_safety_checker
        Added value: +{
        +  "type": "boolean"
        +}
      • addedInput schema / properties / first_frame_image_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / negative_prompt
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / prompt
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / ratio
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / reference_audio_url
        Added value: +{
        +  "type": "string"
        +}
      • addedInput schema / properties / reference_image_urls
        Added value: +{
        +  "items": {},
        +  "type": "array"
        +}
      • addedInput schema / properties / reference_video_urls
        Added value: +{
        +  "items": {},
        +  "type": "array"
        +}
      • addedInput schema / properties / seed
        Added value: +{
        +  "type": "number"
        +}
      • addedInput schema / properties / watermark
        Added value: +{
        +  "type": "boolean"
        +}
  3. 9 tool updatesv0.1.6
    • First observedanimate
    • First observedcheck_pricing
    • First observededit_video
    • First observedget_task
    • First observedimage_to_video
    • First observedlogin
    • First observedspeech_to_video
    • First observedtext_to_image
    • First observedtext_to_video

TDQS

B3.3/5.0

Scored across 9 tools

Disambiguation4/5

Most tools are clearly differentiated by input modality (text_to_image, image_to_video, speech_to_video, etc.), but 'animate' and 'edit_video' have overlapping video-manipulation purposes and vague descriptions, making selection uncertain. Overall, the set is mostly distinct with only minor ambiguity.

Naming Consistency3/5

All names use snake_case, which is consistent, but the patterns vary widely: some follow input_to_output form (text_to_image), some are verb_noun (edit_video, get_task), and some are single verbs (animate, login). This mixed convention reduces predictability.

Tool Count5/5

Nine tools is a well-scoped set for a generation API, covering multiple creation endpoints plus utility functions. No tool feels redundant.

Completeness4/5

The surface covers task creation for various modalities, task status retrieval, pricing, and authentication. However, it lacks task management operations like cancel, list, or delete, which could be needed for longer workflows.

Maintenance

ActivityMaintained
ResponsivenessNo issues

Related MCP Connectors

Related MCP Servers