Holo3: Foundational Models for Navigation and Computer Use Agents Model Description Holo3 is our latest generation of large scale Vision Language Models (VLMs) specifically optimized for GUI Agents . Like its predecessors, it operates across diverse digital environments—web, desktop, and mobile—by interpreting visual interfaces, reasoning over complex content, and executing precise actions. Holo3 achieves state of the art performance on OSWorld Verified , setting a new benchmark for computer use agents. While it retains the world class web navigation capabilities of Holo2 , the new Holo3 35B A3B architecture is designed to thrive in realistic business environments. Developed by: H Company Model type: Vision Language Model for Navigation and Computer Use Agents Architecture: Sparse Mixture of Experts (MoE) with 35B total / 3B active parameters Fine tuned from model: Qwen/Qwen3.5 35B A3B Blog Post: hcompany.ai/holo3 Quickstart: hub.hcompany.ai/quickstart License: Apache 2.0 License Get Started Explore our Quickstart guide to learn how to integrate with our inference API. Training Strategy Holo3 35B A3B is based on the Qwen3.5 architecture and has been reinforced to strengthen its cor…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy