The Definitive Guide toAI Data Centers
Ask the GuideAboutAccount
Guide › Numbers register › Claim

LLMflation: inference cost decline at fixed quality — Epoch AI puts it at ~9–900x/yr depending on the capability threshold, ~40x/yr for GPT-4-level GPQA Diamond

~9–900x/yr by thresholdobserved

Value kindobserved — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range.
ScopeThe headline rate depends entirely on the basis. a16z's original 'LLMflation' series tracks GPT-3-level quality and gives ~10x/yr (~1,000x over three years, ~$60 to ~$0.06 per million tokens); Epoch AI's cross-benchmark median is nearer ~50x/yr, ~40x/yr at the GPT-4-level GPQA Diamond threshold, and ~9x to ~900x/yr across thresholds. Always state which basis a quoted rate uses — these are not restatements of one measurement.
As of2024-2025
SourceEpoch AI, LLM inference price trends
Reviewchecking…review by 2026-09-22 · fast cadence
Recorded changeslast 2026-08-31 · 2 revisions tracked
Claim idllmflation-inference-cost-decline-at-fixed

Where the guide uses it

← Full numbers register — every date-stamped figure in the guide, with revision history.