Provides camera and vision tools for AI assistants to list available cameras, capture images from USB cameras, and save frames to disk for use with LLMs.
Enables PTZ camera control with gimbal positioning, snapshots, and AI visual analysis for OBSBOT and UVC cameras. Supports autonomous scanning patterns and integrates with vision-language models for real-time camera analysis.
Enables LLMs to inspect a live BabylonJS scene through structured scene data, WebGL frame captures, screenshots, and arbitrary scene evaluation, with optional Gemini vision analysis.
An MCP server that enables interaction with local camera devices to capture and process images. It allows LLMs to access video devices with configurable settings such as resolution, orientation, and image format.
Enables capturing JPEG stills from a Raspberry Pi CSI camera and returning them as image content through the capture_image MCP tool for LLM processing.