Vera Rubin NVL72 inference cost per million tokens vs GB200 NVL72 — one-tenth (ratio only; NVIDIA/CoreWeave publish no absolute $/M-token)
~1/10 (one-tenth)observed
| Value kind | observed — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range. |
|---|---|
| Caveat | Relative figure only; any absolute ~$/M-token appears solely in secondary coverage, not in NVIDIA/CoreWeave primaries. |
| As of | 2026-07 |
| Source | NVIDIA (Vera Rubin blog) / CoreWeave — ratio only |
| Review | checking…review by 2026-09-06 · fast cadence |
| Recorded changes | last 2026-07-22 |
| Claim id | vera-rubin-nvl72-inference-cost-per-million-tokens-vs-gb200 |
Where the guide uses it
Not quoted in a chapter yet; it is kept in the curated register.
← Full numbers register — every date-stamped figure in the guide, with revision history.