macOS OCR MCP
Provides OCR (Optical Character Recognition) capabilities using macOS's built-in Vision framework, allowing extraction of text from images with confidence scores and bounding boxes.
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@macOS OCR MCPExtract text from ~/Desktop/screenshot.png"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
macOS OCR MCP Tool
This project provides a MetaCall Protocol (MCP) tool to perform Optical Character Recognition (OCR) on images using macOS's built-in Vision framework. It exposes an ocr_image tool that takes an image file path and returns the recognized text along with confidence scores and bounding boxes.
Project Setup
Dependencies
This project relies on Python 3.13+ and the following main dependencies:
ocrmac: For accessing macOS OCR capabilities. See ocrmac.Pillow: For image manipulation.mcp[cli]>=1.7.1: For the MetaCall Protocol server and client.
Installation
It is recommended to use a virtual environment.
Create and activate a virtual environment:
python -m venv .venv source .venv/bin/activateInstall dependencies using
uv:uv sync
Related MCP server: macOS Native OCR MCP Server
Running the MCP Server
To start the MCP server, run main.py:
uv run main.pyThis will start the MCP server, making the ocr_image tool available.
Available MCP Tools
ocr_image
Description: Conducts OCR on the provided image file using macOS's built-in capabilities. Returns recognized text segments, their confidence scores, and bounding box coordinates.
Input:
file_path: str- The absolute or relative path to the image file.Output (Example Success):
{ "filename": "path/to/your/image.png", "annotations": [ { "text": "Hello World", "confidence": 0.95, "bounding_box": [0.1, 0.1, 0.5, 0.05] }, // ... more annotations ] }Output (Example Error):
{ "error": "OCR functionality is only available on macOS." }or
{ "error": "File not found: path/to/nonexistent/image.png" }
Note: This tool will only function correctly on a macOS system due to its reliance on the Vision framework.
Testing with MCP Inspector
You can use the MCP Inspector to connect to the running MCP server and test the tool.
Cursor MCP Configuration
To configure this MCP server in Cursor, you can add the following to your MCP JSON configuration file (e.g., ~/.cursor/mcp.json or project-specific .cursor/mcp.json):
{
"mcpServers": {
"ocrmac": {
"command": "uv",
"args": [
"--directory",
"/path/to/macos-ocr-mcp",
"run",
"main.py"
]
}
}
}This configuration tells Cursor how to start your MCP server. You can then call the ocrmac.ocr_image tool from within Cursor.
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- AlicenseAqualityCmaintenanceProvides screenshot and OCR capabilities for macOS.Last updated110923MIT
- Flicense-quality-maintenanceProvides offline, high-accuracy OCR capabilities for images and PDFs using macOS's built-in Vision framework. Supports multi-language text extraction with intelligent block aggregation for tables and paragraphs, outputting structured JSON data suitable for document reconstruction.Last updated1
- AlicenseAqualityDmaintenanceEnables AI agents to capture and analyze screenshots of macOS applications, windows, or the entire screen using local (Ollama) or cloud-based AI vision models, with non-intrusive, fast screen capture via Apple's ScreenCaptureKit.Last updated3122MIT
- Alicense-quality-maintenanceA macOS-based MCP server that enables high-accuracy text extraction from PDF and image files using the OwlOCR app or Apple's Vision Framework. It supports multi-language OCR and provides asynchronous tools for processing documents directly within MCP clients.Last updated
Related MCP Connectors
OCR.space MCP — wraps the OCR.space API (ocr.space) for image/PDF → text OCR.
OCR, transcription, file extraction, and image generation for AI agents via MCP.
Turn any PDF into structured JSON via AI + OCR: invoices, bank statements, contracts.
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/whiteking64/macos-ocr-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server