image-analysis-mcp
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Features and capabilities supported by this server
Protocol revision2025-11-25
| Capability | Details |
|---|---|
| tools | {
"listChanged": false
} |
| prompts | {
"listChanged": false
} |
| resources | {
"subscribe": false,
"listChanged": false
} |
| experimental | {} |
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
| ocr_imageA | OCR text from an image. Extracts text using the best available OCR backend. Returns text blocks with confidence scores and bounding boxes. Args: image_path: Absolute path to the image file (must be under home directory) |
| image_metadataA | Extract full metadata from an image. Returns file info (size, timestamps, SHA-256), image properties (format, dimensions, DPI, colour space), and EXIF data including camera make/model, datetime, and GPS coordinates. Args: path: Absolute path to the image file (must be under home directory) |
| extract_textA | Extract all text and metadata from an image. Convenience wrapper that combines OCR text extraction with full image metadata in a single JSON response. Args: image_path: Absolute path to the image file (must be under home directory) |
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
TDQS
Scored across 3 tools
ocr_image and image_metadata have clearly distinct purposes, but extract_text is a combined wrapper that overlaps with both, creating potential confusion about which tool to use for a given task. The descriptions help clarify that extract_text is a convenience option, but the presence of a full-coverage tool makes the specific tools somewhat redundant.
Tool names are all snake_case, but the patterns vary: ocr_image uses an acronym as a verb, image_metadata is a noun-noun compound with no action verb, and extract_text is a clear verb-noun pair. While still readable, this mix prevents a predictable verb_noun convention across the set.
Three tools is on the lower end of typical server scope, but is appropriate for a focused image analysis tool that covers OCR and metadata extraction. The count is not excessive, and each tool fills a specific need, though the combined wrapper could be seen as unessential.
For the apparent domain of text extraction and metadata retrieval, the set is fairly complete: ocr_image covers text, image_metadata covers metadata, and extract_text provides a combined result. However, other common image analysis operations (e.g., object detection, format conversion) are absent, though they may be out of scope for this server.