Muse Glimmer 30B is a dense Apache 2.0 model with image input. VRAM at bf16, 8-bit and 4-bit, which single cloud GPU fits it, and the hourly price.
Blog
Model deployment guides and product updates. We measure every performance number we publish on a real deployment in a real cloud account.
-
Muse Glimmer 30B hardware requirements: what one GPU needs2026-08-28
-
GLM-5.3-Flash (Ox Alpha) hardware requirements and license2026-08-27
Ox Alpha was GLM-5.3-Flash. Z.ai released the weights under MIT on 26 August 2026. The 320B/18B MoE, the 306 GiB FP8 checkpoint, a measured 34-minute cold boot on 4x H200, $/hour.
-
Is Ox Alpha open weights? Yes: it is GLM-5.3-Flash, MIT2026-08-25
Yes, since 26 August 2026. Z.ai confirmed Ox Alpha was GLM-5.3-Flash and published the weights under MIT. What it is, what it runs on, what the clues got right.
-
GLM-5.3 hardware requirements and open-weights release date2026-08-21 · pre-release
Z.ai says the weights follow about two weeks after the August 14 launch. What is public so far, when to expect the weights, and the hardware you will likely need.
-
DeepSeek V4 Flash requirements: 2x H200, boot time, $/hour2026-08-21
GPU requirements, measured cold boot time, and what it costs to run DeepSeek V4 Flash privately in your own cloud account.