Nvidia's Nemotron beat the top human at IOI 2026. The paper's own numbers show the model answering once scored 304, and what the rest of the score cost in GPUs.
Blog
Model deployment guides and product updates. We measure every performance number we publish on a real deployment in a real cloud account.
-
Nemotron's IOI 2026 gold: 760 GPUs, 1,000 tries2026-09-03
-
Qwen3.8-27B hardware requirements: what one cloud GPU needs2026-09-01
Qwen3.8-27B is a dense 27B Apache 2.0 model. What it needs in VRAM at bf16, 8-bit and 4-bit, which single cloud GPU fits it, and what that GPU costs per hour.
-
Muse Glimmer 30B hardware requirements: what one GPU needs2026-08-28
Muse Glimmer 30B is a dense Apache 2.0 model with image input. VRAM at bf16, 8-bit and 4-bit, which single cloud GPU fits it, and the hourly price.
-
GLM-5.3-Flash (Ox Alpha) hardware requirements and license2026-08-27
Ox Alpha was GLM-5.3-Flash. Z.ai released the weights under MIT on 26 August 2026. The 320B/18B MoE, the 306 GiB FP8 checkpoint, a measured 34-minute cold boot on 4x H200, $/hour.
-
Is Ox Alpha open weights? Yes: it is GLM-5.3-Flash, MIT2026-08-25
Yes, since 26 August 2026. Z.ai confirmed Ox Alpha was GLM-5.3-Flash and published the weights under MIT. What it is, what it runs on, what the clues got right.
-
GLM-5.3 hardware requirements and open-weights release date2026-08-21 · pre-release
Z.ai says the weights follow about two weeks after the August 14 launch. What is public so far, when to expect the weights, and the hardware you will likely need.
-
DeepSeek V4 Flash requirements: 2x H200, boot time, $/hour2026-08-21
GPU requirements, measured cold boot time, and what it costs to run DeepSeek V4 Flash privately in your own cloud account.