Skip to main content
Glama
97,196 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Information about Archon" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    A
    quality
    B
    maintenance
    Enables AI agents to drive 360° cameras by connecting, checking status, reading and changing settings, taking photos, recording video, browsing the gallery, and downloading media as cancellable background jobs, with local stitching and export of 360° footage. It is honest about backend limits, returning structured, explained errors for anything a given camera or connection cannot do.
    1
    MIT
  • F
    license
    A
    quality
    D
    maintenance
    Enables AI image generation, editing, and composition using Google's Gemini image models (Nano Banana Pro and Nano Banana). Supports text-to-image generation, multi-image composition, flexible aspect ratios, high-resolution output up to 4K, and real-time information grounding.
    4
    -
  • A
    license
    A
    quality
    A
    maintenance
    MCP server that captures short screen recordings as numbered, timestamped still frames and composes them into a single contact sheet returned as image content, letting AI agents reason about motion, animations, and UI transitions instead of inspecting video frame by frame. Also supports single screenshots and listing targetable windows.
    1
    6
    147 npm
    MIT
  • F
    license
    A
    quality
    C
    maintenance
    Enables AI assistants to recognize and extract information from images via GLM-4V, supporting automatic screenshot recognition and MCP-based local image file reading for non-vision models like DeepSeek.
    1
    -
  • A
    license
    A
    quality
    B
    maintenance
    Enables coding agents to examine a rendered web page by cross-checking what the browser reports about the DOM/CSS against what is actually painted in pixels, pinpointing which element is broken and the evidence why. Reports visual problems without requiring a baseline, and is designed to grow into design-token extraction and accessibility/UX audits built on Playwright.
    1
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Enables web content extraction, screenshot capture, web search, arXiv paper search, and image search through Jina AI's APIs. Provides tools for reading URLs as markdown, searching the web for current information, and finding academic papers or images.
    19
    869
    Apache 2.0
  • A
    license
    A
    quality
    D
    maintenance
    Analyzes webpage screenshots to extract precise layout information by locating image assets and calculating spatial relationships, enabling AI assistants to accurately recreate layouts with proper semantic structure using computer vision.
    3
    6
    MIT
  • A
    license
    A
    quality
    Not graded
    maintenance
    Enables web searching via SearXNG, page content extraction with Crawl4AI, and image analysis using vision language models. It provides AI agents with tools for information synthesis and web-based data retrieval through OpenAI-compatible LLM endpoints.
    3
    Apache 2.0
  • A
    license
    A
    quality
    D
    maintenance
    A universal vision MCP server that enables Claude Code and Claude Desktop to describe images, extract text, and answer questions about images by converting visual content to text via multiple AI providers.
    3
    7 npm
    MIT
  • -
    license
    A
    quality
    Not graded
    maintenance
    Enables comprehensive video file analysis including extracting metadata, stream information, bitrate calculations, and generating technical reports. Supports all FFmpeg-compatible video formats with output in JSON, text, or Markdown formats.
    4
    9 npm
    1
    -
  • A
    license
    B
    quality
    B
    maintenance
    Enables Claude Code and Gemini CLI to see your macOS screen, answering questions about the file, error, UI, database table, or web page you're looking at. Captures screenshots, reads database schemas, and controls Chrome/Brave, presenting answers as on-screen highlights or spoken bubbles above your cursor.
    17
    MIT