What this section answers
The answer is a utilization threshold, not a preference, and the threshold moves with whether you are adding GPUs to existing infrastructure or buying loaded nodes.
decision · section 05 of 10
Rent, reserve, or own — at what utilization and under whose procurement scope?
What this section answers
The answer is a utilization threshold, not a preference, and the threshold moves with whether you are adding GPUs to existing infrastructure or buying loaded nodes.
Boundary
Compliance, capacity certainty, and latency constraints can make the cost-optimal answer the wrong one; the control premium runs roughly 2–4x.
Source coordinate
Report heading If you run ML infrastructure, under What this means for decisions, in The real cost of AI: August 2026.
The GPU market is split: neoclouds at $3-4/GPU-hour, hyperscalers at $7-12. If you’re paying hyperscaler on-demand for sustained inference, you are likely overpaying by 2x. Whether owning beats renting now has a two-part answer. Procurement scope first: a lean team adding GPUs to existing infrastructure (~$30K/GPU) breaks even vs hyperscaler on-demand at ~21% utilization and beats median neocloud pricing above ~37% — but an enterprise node-loaded buyer (~$94K/GPU all-in) needs ~57% just against hyperscaler on-demand, essentially never beats median neoclouds, and never beats cheap neoclouds or spot. Utilization second: above ~70% sustained, on-prem wins on raw cost in the lean regime. And if you’re buying for sovereignty, compliance, or guaranteed capacity, you’re paying a 2–4x control premium on purpose — price it as insurance, not as infrastructure.
The thresholds arranged as the decision you actually have to make.