Cost derivation
All figures derived, none quotedThe price is low for one structural reason: we own the accelerators instead of renting them, at $1.53 a GPU hour against $2.49 to rent the same board. Below is the whole calculation, not a summary of one — change an assumption and every figure on this site moves with it.
Calculation note 02.1 — cost of one GPU hour
The same accelerator rents for about $2.49 an hour. Renting is not a small premium on a large number — it is the number, marked up. That margin is the one thing a better base model cannot take away from us.
Calculation note 02.2 — Qwen3-Coder-480B-A35B at 4 GPU × 3,200 tok/s
Output price against published closed-model list
14× below frontier list for the same unit of output, on a model with comparable coding ability. The cheapest generative tier goes further — Devstral-Small lists at $0.45 per million. Closed list prices are market observations, not our figures; ours are derived above and move when the assumptions move.