The Definitive Guide toAI Data Centers
Ask the GuideAboutAccount
Guide › Numbers register › Claim

Vera Rubin NVL72 inference cost per million tokens vs GB200 NVL72 — one-tenth (ratio only; NVIDIA/CoreWeave publish no absolute $/M-token)

~1/10 (one-tenth)observed

Value kindobserved — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range.
CaveatRelative figure only; any absolute ~$/M-token appears solely in secondary coverage, not in NVIDIA/CoreWeave primaries.
As of2026-07
SourceNVIDIA (Vera Rubin blog) / CoreWeave — ratio only
Reviewchecking…review by 2026-09-06 · fast cadence
Recorded changeslast 2026-07-22
Claim idvera-rubin-nvl72-inference-cost-per-million-tokens-vs-gb200

Where the guide uses it

Not quoted in a chapter yet; it is kept in the curated register.

← Full numbers register — every date-stamped figure in the guide, with revision history.