The Definitive Guide toAI Data Centers
Ask the GuideAboutAccount
Guide › Numbers register › Claim

power-oversubscription headroom: inference (uncorrelated peaks) vs synchronous training

~21% vs ~3%observed

Value kindobserved — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range.
ScopeMeasured on the production LLM fleets and power-management platform characterized by POLCA (Patel et al., ASPLOS 2024). It is not a deployable allowance for a 2026 reasoning, MoE, or disaggregated fleet: derive that from the proposed fleet's measured coincident demand, its tested protection and capping response, and the serving degradation you will allow.
As of2024
SourcePatel et al., POLCA (Microsoft Research), ASPLOS 2024 · Abstract: training clusters offer about 3% headroom, versus about 21% for inference clusters
Reviewchecking…review by 2026-11-23 · standard cadence
Recorded changeslast 2026-07-03 · 2 revisions tracked
Claim idpower-oversubscription-headroom-inference-2

Where the guide uses it

← Full numbers register — every date-stamped figure in the guide, with revision history.