generate_diagram
Convert text prompts into visual architecture diagrams. Select diagram type, theme, resolution, and more.
Instructions
Generate an architecture diagram from a text prompt.
Args:
prompt: Description of what to generate
diagram_type: Type of diagram (architecture|data_flow|component|sequence|integration|infographic|c4_container|exec_infographic|generic)
theme: Background theme (light|dark). Default: light — the portfolio-wide
default (rep / marketing / CISO-facing output is the common case). Pass
theme="dark" for a dark charcoal canvas. This single switch governs the
background; you do NOT need to describe a background color in prompt.
provider: LEAVE AS DEFAULT ("auto"). The server is responsible for picking the right provider and model for the diagram type — callers should describe what to draw and let the server decide how. The current default chain is OpenAI gpt-image-2 (primary) → Gemini (fallback). Override only if you have a specific provider/model comparison need.
model: LEAVE UNSET unless you're explicitly benchmarking models. Server picks the right model for the chosen provider.
resolution: Output resolution (1K|2K|4K)
aspect_ratio: Output aspect ratio (16:9|1:1|9:16|4:3)
style_reference: Style name or path to reference image
output_path: Where to save the image (auto-generated if not provided)
temperature: Generation creativity (0.0 to 2.0)
quality: Output quality tier for OpenAI gpt-image-2 / gpt-image-1-mini (low|medium|high|auto).
Cost scales dramatically: at 1536x1024 on gpt-image-2, low=$0.005, medium=$0.041, high=$0.165.
Ignored by Gemini and legacy gpt-image-1.5. Default: auto.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| model | No | ||
| theme | No | light | |
| prompt | Yes | ||
| quality | No | auto | |
| provider | No | auto | |
| resolution | No | 2K | |
| output_path | No | ||
| temperature | No | ||
| aspect_ratio | No | 16:9 | |
| diagram_type | No | generic | |
| style_reference | No |