Ship a Full-Stack App with One Prompt

Copy this prompt into your AI coding agent, or open it in one below.

Give this to your AI Create a to-do list app using Puter.js

Coding manually? see the guide

hcompany: Holo4 35B A3B

hcompany/holo4-35b-a3b

Try Holo4 35B A3B for free in your browser, and add it to your app for free with Puter.js AI API.

Try it free Add to your app
// npm install @heyputer/puter.js
import { puter } from '@heyputer/puter.js';

puter.ai.chat("Explain quantum computing in simple terms", {
    model: "hcompany/holo4-35b-a3b"
}).then(response => {
    document.body.innerHTML = response.message.content;
});
<html>
<body>
    <script src="https://js.puter.com/v2/"></script>
    <script>
        puter.ai.chat("Explain quantum computing in simple terms", {
            model: "hcompany/holo4-35b-a3b"
        }).then(response => {
            document.body.innerHTML = response.message.content;
        });
    </script>
</body>
</html>

Model Card

Holo4 35B A3B is a vision-language model from H Company built for computer-use agents. It reads screenshots and produces clicks, typing, code, and tool calls.

It is a Mixture-of-Experts model with 35 billion total parameters and 3 billion active per token, based on Alibaba's Qwen3.6-35B-A3B. H Company reports that it improves significantly over its Qwen base models. On the model card, it scores 30.9% on OSWorld 2.0 at $0.61 per task and 34.5% on AutomationBench at $0.02 per task. The larger Holo4 27B variant scores higher on both benchmarks at a higher cost per task on OSWorld 2.0.

It supports a 262,144-token context window and is released under the Apache 2.0 license. It is aimed at developers building agents that operate real graphical interfaces.

Context Window 262K

tokens

Max Output 16K

tokens

Input Cost $0.3

per million tokens

Output Cost $2

per million tokens

Release Date N/A

 

Try Holo4 35B A3B for free

Try Holo4 35B A3B instantly in your browser.
This playground uses the Puter.js AI API — no API keys or setup required.

Chat hcompany/holo4-35b-a3b
Chat with Holo4 35B A3B
Powered by Puter.js

More AI Models From hcompany

Find other hcompany models

Chat

Holo4 27B

Holo4 27B is a dense 27 billion parameter vision-language model from H Company for computer-use agents. It is built on Qwen3.8-27B and takes screenshots and tool results as input, then returns clicks, typing, code, and MCP or API tool calls across desktop, web, and mobile interfaces. H Company trained it with supervised fine-tuning on 127 billion tokens, mostly agent operation data, followed by reinforcement learning. It scores 85.2% on OSWorld, against 84.3% for Qwen3.8-27B, and 61.7% on OSWorld 2.0, against 48.0% for Qwen3.8-27B. On OSWorld 2.0 it trails Opus 5.5, which scores 81.8%. It also scores 45.4% on AutomationBench. It supports a 262,144 token context window and suits developers building agents that operate real software interfaces.

Chat

Holo3.1 35B A3B

Holo3.1 35B A3B is an open-weights vision-language model from H Company for computer-use and GUI agents, operating web, desktop, and mobile interfaces from screenshots. It is a sparse Mixture-of-Experts model with 35 billion total parameters and 3 billion active per token, built on the Qwen 3.5 family. It supports native function calling in addition to structured JSON outputs, which suits agent frameworks that return actions as tool calls. H Company reports 79.3% on AndroidWorld, and says it improves on Holo3 on OSWorld and on its Holotab harness. It was evaluated against Holo3, Qwen3.5, Kimi-K2.5, and Claude Sonnet 4.6. Released June 1, 2026 under an Apache 2.0 license, it is aimed at developers building agents that click, type, and navigate real interfaces.

Chat

Holo3 122B A10B

Holo3 122B A10B is a vision-language model from H Company built for GUI and computer-use agents, reading screenshots and deciding the actions needed to complete tasks across web and desktop applications. It is a sparse Mixture-of-Experts model with 122 billion total parameters and 10 billion active per token, based on Qwen3.5. H Company reports a score of 78.85% on OSWorld-Verified, a benchmark of multi-step desktop tasks, compared with 77.8% for its smaller Holo3 35B A3B model. Released March 31, 2026, it is aimed at developers building agents that click, type, and navigate real interfaces, such as browser automation and business software workflows, rather than general chat or writing tasks.

Frequently Asked Questions

How do I use Holo4 35B A3B?

You can access Holo4 35B A3B by hcompany through Puter.js AI API. Include the library in your web app or Node.js project and start making calls with just a few lines of JavaScript — no backend and no configuration required. You can also use it with Python or cURL via Puter's OpenAI-compatible API.

Can I try Holo4 35B A3B for free?

Holo4 35B A3B is free to try with a Puter account. Every account includes a free AI allowance, and you can chat with it in the playground on this page. You can upgrade your account anytime for a larger allowance.

Is the Holo4 35B A3B API free for developers?

Holo4 35B A3B is free to integrate using the Puter.js AI API. With the User-Pays Model, you can add AI to your app for $0, since users cover their own AI usage through their Puter account.

What is the pricing for Holo4 35B A3B?
Holo4 35B A3B costs $0.3 per 1M input tokens and $2 per 1M output tokens.
Price per 1M tokens
Input$0.3
Output$2
What is the context window of Holo4 35B A3B?

Holo4 35B A3B supports a context window of 262K tokens. For reference, that is roughly equivalent to 524 pages of text.

What is the max output length of Holo4 35B A3B?

Holo4 35B A3B can generate up to 16K tokens in a single response.

Does it work with React / Vue / Vanilla JS / Node / etc.?

Yes — the Holo4 35B A3B API works with any JavaScript framework, Node.js, or plain HTML through Puter.js. Just include the library and start building. See the documentation for more details.

Add Holo4 35B A3B to your app for free

Developers can integrate Holo4 35B A3B for free using the Puter.js AI API.
With the User-Pays Model, each user covers their own AI usage instead of the developer.

Get started How pricing works