Skip to main content
Glama
tilikumotp

Qwen-GroundingDINO-Visual-MCP

by tilikumotp

Related Servers

Alternatives to Qwen-GroundingDINO-Visual-MCP

No user-submitted related servers found.

    Related Servers

    • A
      license
      Not graded
      quality
      B
      maintenance
      MCP server for multimodal understanding and object grounding (bounding boxes) across images, videos, and documents, with support for multiple AI providers (Zhipu GLM-V, OpenAI GPT-4o, Anthropic Claude, or any OpenAI-compatible endpoint).
      1
      MIT
    • A
      license
      Not graded
      quality
      C
      maintenance
      MCP server that provides a 'borrowed eye' for text-only LLMs, enabling them to identify and describe local images via the Qwen VL vision model, including face recognition, scene description, OCR, and targeted visual questioning.
      4 npm
      Apache 2.0
    • A
      license
      Not graded
      quality
      C
      maintenance
      A local MCP server that gives LLMs eyes for images by performing object detection (YOLOv8) and text recognition (EasyOCR), outputting descriptive statements about objects and text positions without any API key or cloud dependency.
      MIT