Skip to main content
Glama
jonchun

Gemini Image Generator MCP Server

by jonchun

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
GEMINI_MODELNoGemini model to usegemini-2.5-flash-image
GEMINI_API_KEYYesYour Gemini API key
GEMINI_BASE_URLNoAPI base URLhttps://generativelanguage.googleapis.com
DEFAULT_OUTPUT_IMAGE_PATHNoDefault save location for imagesCurrent directory

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Features and capabilities supported by this server

Protocol revision2025-11-25

CapabilityDetails
tools
{
  "listChanged": false
}
prompts
{
  "listChanged": false
}
resources
{
  "subscribe": false,
  "listChanged": false
}
experimental
{}

Tools

Functions exposed to the LLM to take actions

NameDescription
generate_image_from_textA

Generate an image from a text prompt using Gemini.

Args: prompt: Text description of the desired image. output_dir: Optional directory to save the generated image. If not provided, the image is only returned in the response (not saved to disk). model: Optional Gemini model name. If not provided, uses GEMINI_MODEL environment variable. ctx: Optional context for progress reporting.

Returns: List containing ImageContent with the generated image, and optionally TextContent with file path if saved.

transform_image_from_encodedA

Transform a base64-encoded image using Gemini.

Args: encoded_image: Base64 data URL (data:image/[format];base64,[data]). prompt: Text description of desired transformation. output_dir: Optional directory to save the generated image. If not provided, the image is only returned in the response (not saved to disk). model: Optional Gemini model name. If not provided, uses GEMINI_MODEL environment variable. ctx: Optional context for progress reporting.

Returns: List containing ImageContent with the transformed image, and optionally TextContent with file path if saved.

transform_image_from_fileA

Transform an image file using Gemini.

Args: image_file_path: Path to the source image file. prompt: Text description of desired transformation. output_dir: Optional directory to save the generated image. If not provided, the image is only returned in the response (not saved to disk). model: Optional Gemini model name. If not provided, uses GEMINI_MODEL environment variable. ctx: Optional context for progress reporting.

Returns: List containing ImageContent with the transformed image, and optionally TextContent with file path if saved.

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources

TDQS

A4/5.0

Scored across 3 tools

Disambiguation5/5

Each tool has a clearly distinct purpose: text-to-image generation versus image transformation from two different input sources (encoded data or file path). The input types prevent confusion, and descriptions explicitly state the expected arguments.

Naming Consistency5/5

All tool names follow the same snake_case verb_noun_preposition pattern (generate_image_from_text, transform_image_from_encoded, transform_image_from_file). The convention is consistent and predictable.

Tool Count5/5

Three tools are well-scoped for an image generation and transformation server; each tool covers a distinct input method without redundancy. No tool feels superfluous or missing from a minimal set.

Completeness4/5

The server covers the core lifecycle: generating an image from text and transforming existing images from both encoded data and file paths. Minor gaps exist (e.g., no batch generation or targeted editing), but the primary workflows are supported.

Maintenance

ActivityInactive
ResponsivenessNo issues