generate_image
Generate image files from text prompts with configurable aspect ratio, style, and brand colors, and receive exact absolute paths for direct use.
Instructions
Generate image(s) from a text prompt at the given aspect ratio (16:9, 1:1, 3:4, 4:3, 9:16, or WxH). Generates ONCE and returns the exact ABSOLUTE file path(s) saved, e.g. {"paths": ["C:/.../img-....png"]}. Callers should use the returned path directly and never re-generate to "find" the file.
Images are saved into out_dir (created if missing) as img--.png.
When enhance is True (default), the prompt is auto-expanded via the ChatGPT text path before drawing. style='slide' = clean editorial look; style='fintech' = light-blue dashboard look; style='auto' is the general default.
thinking sets reasoning effort: 'auto' (ChatGPT default) or 'standard'/ 'extended'/'max' (increasing). Higher effort improves rendered-text fidelity (e.g. Vietnamese diacritics) at the cost of speed.
brand_colors (list of hex like ['#10B981']) forces a palette; reserve_corner (e.g. 'top-left') keeps a corner clear for a logo and bans model-drawn logos/text. With enhance=False these still apply via the offline template.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| n | No | ||
| style | No | auto | |
| aspect | No | 16:9 | |
| prompt | Yes | ||
| enhance | No | ||
| out_dir | No | out | |
| thinking | No | auto | |
| brand_colors | No | ||
| reserve_corner | No |