Enables LLMs like DeepSeek to understand images by calling external vision models via OpenAI-compatible API. Provides tools to describe images or diagnose connectivity.
Provides a vision tool that converts images into structured text descriptions and OCR using the free GLM-4.6V-Flash model, enabling text-only LLMs like DeepSeek to understand images.