Skip to main content
Glama
530,137 tools. Updated 2026-09-07 19:11

"Tools for OCR and Analyzing Text and Content in Images" matching MCP tools:

  • Search past screen capture OCR text to find when you saw specific content, error messages, or work references. Returns matching entries with timestamps and context.
    AGPL 3.0
  • Extract visible text from your screen using OCR. Returns text grouped by detected windows with bounding boxes. Use when you need code, terminal, chat, or document content without visual layout. Avoids sending data to the cloud.
    AGPL 3.0
  • Batch-analyze multiple videos in one call, extracting transcripts, frames, OCR text, and metadata per video with configurable concurrency.
    MIT
  • Extract transcript, key frames, OCR text, and metadata from any video URL or local file. Provides structured analysis for AI agents.
    MIT

Matching MCP Servers

  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables MCP clients to connect to a privacy-first, self-hostable workout planning and training log, allowing coaching agents to preview and apply program changes while accessing training data through OAuth-protected endpoints.
    AGPL 3.0
  • A
    license
    Not graded
    quality
    C
    maintenance
    Enables AI assistants to interact with Databricks workspaces, running SQL queries, managing jobs, and exploring schemas via the Model Context Protocol.
    1
    GPL 3.0

Matching MCP Connectors

  • Rick and Morty MCP — wraps the Rick and Morty API (free, no auth)

  • Exact character/word counting, reversal, palindrome checks, indexing, sorting; Unicode-safe.

  • Extract a webpage's structure—title, headings, links, images, and text—with a static HTML fetch. Get clean content without JavaScript rendering.
    MIT
  • Analyze an image URL with Google Lens to find visual matches, detect objects, extract text via OCR, and identify exact product matches.
    MIT
  • Extract frames, OCR text, and transcript snippets from a specific video time range. Merge visual and audio content into a unified, annotated timeline.
    MIT
  • Extract all visible text from a device screen via OCR, solving cases where UI elements render as images or canvas. Returns text with bounding boxes, sorted top-to-bottom.
    MIT
  • Specify a screen region with device pixel coordinates to extract text via OCR, returning JSON with text, confidence, and bounding box for targeted analysis.
    MIT
  • Extract text from local image files using OCR. Converts image-based content into structured JSON and a text summary for downstream processing.
    MIT
  • Verify identity documents by submitting front and optional back images. Get structured OCR data and authenticity checks for fraud prevention.
    MIT
  • Convert scanned PDFs or images into searchable PDFs by running OCR and adding an invisible text layer, enabling full-text search in documents that were previously unsearchable.
    MIT
  • Force OCR text extraction on scanned PDFs when normal text extraction returns garbled or empty text.
    MIT
    Destructive
  • Read text content from PDF, TXT, MD, DOCX, or CSV files. Supports page ranges and auto-OCR for scanned PDFs.
    MIT
    Destructive
  • Detect PII, PHI, PCI, and secrets in local files—including images, PDFs, and scans—using OCR. Returns found data element types without modifying the file.
    MIT