forge-mcp
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
| FORGE_URL | No | Base URL of the Forge/A1111 --api server | http://127.0.0.1:7860 |
| FORGE_MCP_TOKEN | No | If set, require Authorization: Bearer <token> on the HTTP transport | |
| FORGE_GEN_TIMEOUT | No | Seconds to wait on Forge (covers a cold model-load + generation) | 180 |
| FORGE_DEFAULT_MODEL | No | Checkpoint used when a request doesn't name one | DreamShaperXL_Turbo_SFW |
| FORGE_PREVIEW_MAX_PX | No | Longest edge of the inline JPEG preview | 768 |
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Features and capabilities supported by this server
Protocol revision2025-11-25
| Capability | Details |
|---|---|
| tools | {
"listChanged": false
} |
| prompts | {
"listChanged": false
} |
| resources | {
"subscribe": false,
"listChanged": false
} |
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
| generate_imageA | Generate an image (txt2img) on the local SDXL box. Returns an inline preview + (in the text block) the exact params and the full-res PNG base64 to save in your repo. |
| edit_imageA | Edit/vary an image (img2img). ⚠ HOLISTIC — re-derives the WHOLE frame from the init, so use it for
variants that should re-settle (wardrobe, pose, lighting, hair-mass), NOT single-feature fixes (those
belong in a PSD — pushing just the eyes also moves skin/age/expression). For character consistency,
pass the character's canonical keeper as |
| list_modelsA | List the image models (checkpoints) available on the Forge box, and which one is loaded. |
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
TDQS
Scored across 3 tools
generate_image (txt2img), edit_image (img2img), and list_models each have a clearly distinct purpose, and the descriptions explicitly explain the txt2img vs img2img boundary (including when to use PSD instead). No meaningful overlap remains.
All three tools follow a clean verb_noun snake_case pattern: generate_image, edit_image, list_models. Fully predictable and consistent.
Three tools is a coherent, well-scoped set for a focused SDXL render server, but it sits at the thin end—common needs like upscaling or inpainting have no dedicated tool.
Core lifecycle (text-to-image, image-to-image, model discovery) is covered and the model arg handles checkpoint switching. Minor gaps like upscaling/detail-fix or masked inpainting are acknowledged as out of scope but would otherwise cause dead ends.