Enables local AI image generation on Apple Silicon Macs using MLX and Stable Diffusion. Supports conversational design iteration, asset generation, and wireframe creation with zero API costs through the Model Context Protocol.
Enables text-only LLMs to perceive images entirely on-device, providing vision capabilities like image description, OCR, table extraction, UI analysis, and region focusing without any cloud APIs or API keys.
Enables image and video generation across GPT-Image, Gemini, Grok, and Jimeng with file-based outputs, multi-reference support, and model capability lookup.