Holo4: open-weight computer-use agents for screen, code, and API

H Company's 27B dense and 35B MoE drive desktop, web and Android — ~79% fewer tokens than their base, with open weights and GGUF quants out today.

Nowline Sep 29 1:00 AM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • One model, any interface

    Holo4 clicks and types on screens, writes and runs its own code, and calls MCP or APIs — using whichever fits the task across desktop, web, Android, and code sandboxes. Two open sizes ship today: a 27B dense model and a 35B-A3B MoE. What you can build with it: self-hosted back-office automation, cross-platform QA bots, or internal tools that need auditable action logs.

  • 79% fewer tokens than its base

    On a Godot "build Pac-Man" task, Holo4-27B finished in 2.4M tokens over 68 tool calls versus 11.4M tokens and 197 calls for its base model (Qwen 3.8 27B) — about 79% fewer tokens and 65% fewer tool calls. For long agent loops, that gap is the difference between a run you can afford and one you can't.

  • The OSWorld gap, stated plainly

    It isn't frontier: Holo4-27B scores 61.7% on OSWorld 2.0 against Opus 5.5's 81.8%. The trade is openness — weights you can self-host, cheaply, with full control of the action log, for teams that can't send their screens to a closed API.

  • Weights are up, in every quant

    Grab it now on Hugging Face in BF16, FP8, NVFP4, and 4-bit GGUF — the 4-bit 27B is small enough to run locally. Commercial use is allowed through the managed H Models API or self-hosting; check the model card for exact license terms before you ship.