Private desktop agents
Run the model and agent loop on local hardware so screenshots and actions stay within the user's network.
Independent tool overview
Holo3.1 is H Company's open family of vision-language models for computer-use agents, spanning 0.8B to 35B-A3B parameters with support for web, desktop, mobile, function calling, cloud APIs, and fully local deployment.
Visit the official Holo3.1 site ↗
Overview
Holo3.1 is built for agents that operate graphical interfaces by looking at screenshots and deciding what to click, type, or do next. It targets browser automation, desktop software, Android workflows, and multi-application business tasks rather than general chat.
The family includes 0.8B, 4B, 9B, and 35B-A3B models based on Qwen 3.5. The smaller checkpoints trade accuracy for lower hardware cost, while the mixture-of-experts 35B model activates roughly 3B parameters per step and is H Company's strongest open Holo3.1 option.
Holo3.1 can be self-hosted under Apache 2.0 or accessed through H Company's OpenAI-compatible API. Quantized 35B checkpoints in FP8, NVFP4, and Q4 GGUF make private local execution possible on supported NVIDIA and Apple hardware.
Use cases
The strongest fit depends on the job you need the product to complete, not the size of its feature list.
Run the model and agent loop on local hardware so screenshots and actions stay within the user's network.
Build one computer-use stack for websites, native desktop apps, Android interfaces, and multi-app workflows.
Use smaller models, quantized weights, or the low-cost hosted API for tasks that require many screenshots and steps.
Integrate structured JSON outputs or native function calling into an existing automation harness.
Capabilities
Understands and acts across browser, desktop, and mobile graphical interfaces.
Choose from lightweight 0.8B, 4B, and 9B models or the higher-performing 35B-A3B mixture-of-experts model.
Adds native tool-calling protocols alongside Holo's structured JSON action format.
The 35B release includes FP8, NVFP4, and Q4 GGUF variants for different accelerators and local runtimes.
Use the hosted 35B-A3B model through a familiar text-and-image API without changing the surrounding agent architecture.
H Company's open desktop client can connect to the managed API or a locally served Holo3.1 endpoint.
Process
Step 1
Balance task difficulty, latency, memory, privacy, and hardware constraints before selecting a checkpoint.
Step 2
Use H Company's API for quick testing or serve an Apache-licensed checkpoint with a compatible local runtime.
Step 3
Pair the model with HoloDesktop or another framework that captures screenshots and executes approved actions.
Step 4
Limit accessible apps, accounts, file paths, and high-impact actions, and require confirmation where appropriate.
Step 5
Measure success, latency, retries, cost, and failure recovery on the interfaces your organization actually uses.
Cost
The Holo3.1 model weights are free to download under Apache 2.0; self-hosters pay their own hardware costs. H Company's hosted 35B-A3B API includes a free rate-limited tier and usage-based paid access.
Free
Download and self-host Holo3.1 under Apache 2.0.
$0
Rate-limited API access to Holo3.1 35B-A3B for testing.
$0.25 input / $1.80 output per 1M tokens
Pay-as-you-go access to the Holo3.1 35B-A3B API with higher rate limits.
Pricing checked . Check current pricing at the source ↗
Assessment
Compare
The right alternative depends on the specific output, workflow, controls and budget your project requires.
Agents
For Google's hosted specialized model for agents that interact with interfaces.
Explore Gemini Computer Use →Consumer
For a packaged local-first agent experience designed around personal NVIDIA hardware.
Explore Portable Computer →Agents
For a consumer-oriented local orchestrator across files, native apps, and the web.
Explore Perplexity Personal Computer →Agents
For a desktop product that brings Manus agent workflows onto a user's local machine.
Explore My Computer →Questions
It is a family of vision-language models from H Company designed to power agents that navigate web, desktop, and mobile interfaces.
Yes. H Company publishes open checkpoints, including quantized 35B variants for NVIDIA and Apple hardware, and HoloDesktop can connect to a local OpenAI-compatible server.
H Company publishes the Holo3.1 model family under the Apache 2.0 license, including the hosted 35B-A3B model's weights.
Use 0.8B, 4B, or 9B when hardware and latency matter most; use 35B-A3B when stronger navigation accuracy justifies more memory and compute.
The hosted 35B-A3B API has a free 10-RPM tier. Paid access is currently $0.25 per million input tokens and $1.80 per million output tokens.
No. The model interprets screenshots and produces actions; an agent harness such as HoloDesktop must capture the screen, execute those actions, and enforce permissions.
Bottom line
Holo3.1 is a compelling foundation for teams that want cross-platform computer use without committing to a closed frontier model. Its open license, small-to-large model range, local quantizations, and inexpensive API make experimentation accessible. Production use still depends on a reliable harness, careful task-specific evaluation, and strict approval boundaries around sensitive or irreversible actions.
Visit Holo3.1 website ↗
Gemini Spark - Google's personal agent that runs 24/7 on Cloud VMs

Hermes Desktop - Nous Research's persistent-memory agent as a native desktop app

Higgsfield Supercomputer - ACloud-based AI agent with built-in tools, persistent memory, and scheduled task automation

Kimi Work - Moonshot's local desktop agent running up to 300 parallel agents

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.
Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.
To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.
Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.
Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.