Enables AI agents to operate local desktops and Chromium browsers through MCP tools, unifying accessibility trees, physical input, screenshots, DOM/ARIA, visual grounding, and result verification.
Enables AI hosts to drive cross-platform desktop GUI automation and browser control, providing tools to read, click, type, send shortcuts, screenshot, and verify GUI elements and web pages.