Holotron4-30B-A3B-FP8

Holo4 family:

H Models API Agent trajectories

Model summary

Holotron4-30B-A3B-FP8 is a vision-language model (VLM) for Computer Use, built on NVIDIA Nemotron 3 Nano Omni and developed by H Company. Used with the hai-agents harness, it can send screenshots and tool results to the model, then execute its requested clicks, typing, code, and tool calls.

SpecificationValue
Model IDHcompany/Holotron4-30B-A3B-FP8
ArchitectureNemotronH Nano Omni
Checkpoint formatFP8 safetensors
Context length262,144 tokens

This demo shows Holo4-27B using FreeCAD to build a replica of the Eiffel Tower. More examples in the blog post.

Prompt

Build a 3D Eiffel Tower in FreeCAD at a scale of 1 mm per metre. Center it on the origin and align it with the X and Y axes. Its plan must stay square at every height. The distance from the center to each corner is 62.5 mm at ground level, 32.5 mm at height 57, 17.5 mm at height 115, and 9.35 mm at height 276. Connect these widths with a smooth curve that narrows quickly near the base and more slowly near the top.

Make four separate, identical square legs, one in each quadrant. Their outer corners follow that curve, and each leg narrows from 14 mm across at ground level to 4 mm at height 276. Leave the space between the legs open. Add centered square platforms measuring 72 mm by 72 mm by 4 mm at height 57, 40 mm by 40 mm by 3 mm at height 115, and 22 mm by 22 mm by 3 mm at height 276. Add a square mast from height 276 to 324, tapering from 8 mm across to 2 mm across. Make every component a closed solid with nonzero volume, without filling the space between the legs.

Usage

The harness sends screenshots and tool results to Holotron4, executes the model's requested actions, and sends the results back. It can give the model access to application tools and code execution.

Refer to the documentation for more details about:

Performance

Holotron4 improves over its base model, Nemotron 3 Nano Omni, on GUI workflows and in environments with MCP tools, APIs, or code sandboxes. Gains are absolute percentage points.

BenchmarkInterfaceNemotron 3 Nano OmniHolotron4-30B-A3BGain
OSWorldGUI21.076.3+55.3
OSWorld 2.0GUI and code0.27.9+7.7
AutomationBenchMCP19.435.6+16.2
PinchBenchTerminal84.788.6+3.9
ALE (Linux, code)Terminal0.68.5+7.9

Open-source evaluation traces

For transparency, we share all agent trajectories in the open-source dataset at Hcompany/trajectories.

Training

Holo4 training pipeline

License

The model is governed by the NVIDIA Open Model Agreement.

Downloads last month
136
Safetensors
Model size
33B params
Tensor type
BF16
·
F8_E4M3
·
F32
·
I64
·

Collection including Hcompany/Holotron4-30B-A3B-FP8

Article mentioning Hcompany/Holotron4-30B-A3B-FP8