Smart Image Generator
Generate and edit images using Google's Gemini 2.5 Flash Image with intelligent prompt enhancement and advanced features for professional image creation.
AI-Powered Image Generation: Create high-quality images from text prompts using advanced AI models
Intelligent Prompt Enhancement: Automatically optimize prompts for superior image quality (can be disabled for full control)
Image Editing: Transform existing images with natural language instructions while preserving original style
Multi-Image Blending: Combine multiple visual elements to create composite scenes
Character Consistency: Maintain character appearance across different generations
World Knowledge Integration: Generate accurate context with historical figures, landmarks, and factual scenarios
Flexible Output: Support for PNG, JPEG, and WebP formats with customizable file naming and directory saving
Customizable Parameters: Control aspect ratio and enable/disable specific features as needed
Enables AI-powered image generation and editing using Google's Gemini 2.5 Flash Image API, with automatic prompt enhancement via Gemini 2.0 Flash for superior image quality
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@Smart Image GeneratorGenerate a futuristic cityscape at night with flying cars and neon lights"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
MCP Image Generator ๐
Generate and edit images from Codex, Cursor, Claude Code, or any MCP client. mcp-image adds visual direction to your request before sending it to Gemini, OpenAI, or BytePlus Seedream.
Tell it what image to create or what to change in an existing image, and what it is for. The result is saved to disk and returned to your assistant.
What It Does
Before generating an image, mcp-image rewrites short requests into more specific prompts. It keeps what you asked for and fills in details such as composition, lighting, and camera angle. The more detail you provide, the less it changes.
You ask:
"A photo of a roast chicken dinner for a recipe site. It should look like it was actually cooked, and it should be partway through being carved so you can tell how juicy it is."
mcp-image sends to the image model:
"... a beautifully roasted whole chicken, golden-brown and glistening, resting on a rustic wooden cutting board. One leg is partially carved, revealing tender, succulent white meat and rich, glistening juices pooling around the carving knife ... shallow depth of field focused on the carved chicken."

Generated with Gemini using the default fast quality preset.
What carried through:
for a recipe site: one clear subject, with everything else kept subordinateactually cooked: uneven browning and juices across the boardpartway through being carved: the cut face and slices beside ithow juicy it is: close framing and shallow depth of field around the cut
Baseline from the same request, with prompt enhancement disabled.
Set SKIP_PROMPT_ENHANCEMENT=true to send the original prompt to the image model unchanged.
Related MCP server: Gemini 2.5 Flash Image MCP
Quick Start
You need Node.js 22 or later, an MCP-compatible client, and an API key for one image provider.
1. Get an API key
All three providers generate and edit images. Gemini is the default and requires the least configuration.
Provider | Image size | Output format | Setup |
Gemini (default) | 1K, 2K, 4K | Automatic | Get a key, then set |
OpenAI | 1K, 2K, 4K | PNG or JPEG | Get a key, then set |
BytePlus Seedream | 1K, 2K | PNG or JPEG | Get an AP region key, then set |
Google Search grounding is available with Gemini only. OpenAI may require organization verification before it can generate images.
The examples below use Gemini. Replace the provider settings if you prefer OpenAI or Seedream.
2. Configure your MCP client
Codex
Add this to ~/.codex/config.toml:
[mcp_servers.mcp-image]
command = "npx"
args = ["-y", "mcp-image"]
[mcp_servers.mcp-image.env]
GEMINI_API_KEY = "your_gemini_api_key_here"
IMAGE_OUTPUT_DIR = "/absolute/path/to/images"Cursor
Add this to ~/.cursor/mcp.json for all projects, or .cursor/mcp.json in a project:
{
"mcpServers": {
"mcp-image": {
"command": "npx",
"args": ["-y", "mcp-image"],
"env": {
"GEMINI_API_KEY": "your_gemini_api_key_here",
"IMAGE_OUTPUT_DIR": "/absolute/path/to/images"
}
}
}
}Claude Code
Run this in your project directory:
claude mcp add mcp-image --env GEMINI_API_KEY=your-api-key --env IMAGE_OUTPUT_DIR=/absolute/path/to/images -- npx -y mcp-imageAdd --scope user after mcp-image to make it available in every project.
Never commit API keys to version control. Use an absolute IMAGE_OUTPUT_DIR in MCP configuration because the server's working directory depends on the client. If omitted, images are written to ./output relative to that working directory.
3. Generate an image
Restart your MCP client after changing its configuration, then ask your AI assistant:
Generate a product photo of a ceramic coffee mug on a wooden desk.The generated file is saved in the configured output directory and returned to the assistant as an MCP resource.
pnpm install
pnpm run buildConfigure the MCP client to run the local build instead of npx -y mcp-image:
node /absolute/path/to/mcp-image/dist/index.jsMore Examples
Edit an existing image
Give the assistant an absolute path to the source image:
Edit /path/to/image.jpg so the person is facing right.Control the result
Generate a high-quality product photo of a smartphone with clear text on the screen.Generate a cinematic desert landscape in a 21:9 aspect ratio.Keep the knight's appearance consistent with the previous image.
See the tool reference for the options your assistant can pass explicitly.
Configuration
Changing the provider changes both prompt enhancement and image generation. The way you ask for an image stays the same.
Quality
IMAGE_QUALITY accepts fast (default), balanced, or quality. Set it in the MCP server environment:
IMAGE_QUALITY=balancedA request-level quality option takes precedence. Each provider maps the three values to its own image settings.
Environment variables
Variable | Default | Description |
|
| Default provider: |
| - | API key for Gemini |
| - | API key for OpenAI |
| - | ModelArk AP API key for Seedream |
|
| Directory where generated images are saved; use an absolute path in MCP configuration |
|
| Default quality preset: |
|
| Set to |
You can configure keys for more than one provider and switch per request. A request-level provider option takes precedence over IMAGE_PROVIDER.
Tool Reference
Your MCP client calls this tool for you. Open the reference when you need to check an option or provider limitation.
Parameter | Type | Required | Description |
| string | Yes | Image description or editing instruction |
| string | No |
|
| string | No |
|
| string | No | Absolute path to an input image for editing |
| string | No | Output filename; |
| string | No |
|
| string | No |
|
| boolean | No | Add blending guidance when combining visual elements |
| boolean | No | Keep a character's appearance consistent across images |
| boolean | No | Add context for historical figures, landmarks, and factual scenes |
| boolean | No | Gemini only. Use Google Search grounding for current information |
| string | No | Intended use, such as |
Troubleshooting
API key not found
Check that the key for the selected provider is present in the MCP server's environment:
Gemini:
GEMINI_API_KEYOpenAI:
OPENAI_API_KEYSeedream:
ARK_API_KEY
Restart the MCP client after changing its configuration.
Input image file not found
Use an absolute path and make sure the MCP server can read the file. Input images can be PNG, JPEG, or WebP and must be no larger than 10 MB. Seedream editing accepts PNG and JPEG only.
Provider rejects a request
Check the requested size in the provider table. useGoogleSearch works with Gemini only, and Seedream does not support 4K. For OpenAI permission errors, check your organization settings. For quota or rate-limit errors, check the selected provider account.
Image Generation Prompt Skill
This repository also includes an Agent Skill for assistants that already have access to an image generation tool. It teaches the prompt-writing approach used by mcp-image and works independently of this server.
Install it with:
npx mcp-image skills install --path <skills-directory>For example, use ~/.codex/skills, ~/.cursor/skills, or ~/.claude/skills as the destination.
License
MIT License. See LICENSE for details.
Need help? Open an issue or check Troubleshooting.
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Tools
Related MCP Servers
- AlicenseAqualityDmaintenanceEnables conversational image generation and editing with Google's Gemini 2.5 Flash Image Preview. Supports text-to-image generation, natural language image editing, multi-image composition, and style transfer with optional file saving.4203MIT
- AlicenseNot gradedqualityCmaintenanceUse Nano Banana Pro to generate image from text prompt and edit image3Apache 2.0
- AlicenseAqualityDmaintenanceExposes Google Gemini's Nano Banana image generation models to Claude, enabling text-to-image generation, image editing, and multi-image composition through natural language prompts.3MIT
Related MCP Connectors
Generate images, video & speech with Nano Banana, Veo, Omni and Gemini TTS. Pay as you go.
AI image, video & music generation. Flux, Veo 3.1, Suno V5. Free tier included.
Analyze images from multiple angles to extract detailed insights or quick summaries. Describe visuโฆ
Appeared in Searches
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/shinpr/mcp-image'
If you have feedback or need assistance with the MCP directory API, please join our Discord server