gemini-vision-mcp-safe
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
| HTTP_PROXY | No | HTTP proxy URL. | |
| HTTPS_PROXY | No | HTTPS proxy URL (e.g., http://127.0.0.1:7890). | |
| GEMINI_API_KEY | Yes | Your key from Google AI Studio. | |
| GEMINI_VISION_MODEL | No | Default model for vision requests. | gemini-2.5-flash |
| GEMINI_VISION_ALLOW_URL | No | Allow URL inputs. | true |
| GEMINI_VISION_MAX_IMAGE_MB | No | Maximum image size in MB. | 10 |
| GEMINI_VISION_ALLOW_LOCAL_FILE | No | Allow local file inputs. | true |
| GEMINI_VISION_BLOCK_LOCAL_URLS | No | Block private/loopback IPs for URLs. | true |
| GEMINI_VISION_GEMINI_TIMEOUT_MS | No | Timeout for Gemini API call in milliseconds. | 60000 |
| GEMINI_VISION_REQUEST_TIMEOUT_MS | No | Timeout for URL fetch in milliseconds. | 20000 |
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Features and capabilities supported by this server
Protocol revision2025-11-25
| Capability | Details |
|---|---|
| tools | {
"listChanged": true
} |
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
| analyze_image_with_geminiA | Analyze a local image file or image URL using Google Gemini Vision. PRIVACY RULES:
Supported inputs:
|
| analyze_images_batchA | Batch analyze multiple images (2-5) using Google Gemini Vision. Same privacy rules as analyze_image_with_gemini apply. Useful for comparing screenshots, before/after views, or multi-page documents. Supported inputs: local file paths or HTTP/HTTPS image URLs (can mix). |
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
TDQS
Scored across 2 tools
The two tools have clearly distinct purposes: one for single image analysis and one for batch analysis (2-5 images). There is no overlap, and the descriptions make the difference explicit.
Both tools start with 'analyze' but differ in structure: 'analyze_images_batch' uses plural and 'batch' suffix, while 'analyze_image_with_gemini' uses singular and includes 'with_gemini', which is redundant given the server name. This inconsistency in naming conventions reduces clarity.
With only 2 tools, the set feels thin for a vision analysis server. While these cover basic single and batch analysis, more tools (e.g., different analysis types or output formats) would be expected for a well-scoped server.
The tools provide basic functionality for image analysis (single and batch), but lack features like returning different output formats, handling streaming, or supporting varied prompts. The surface is minimal and may leave agents needing additional capabilities.