The Definitive Guide toAI Data Centers
Ask the GuideAboutAccount
Guide › Numbers register › Claim

Meta Llama 3 405B pre-training: 41% BF16 MFU on 16,384 H100 GPUs at 8,192-token sequence length

~41%observed

Value kindobserved — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range.
ScopeNamed 405B pre-training configuration; other Table 4 layouts report 38% or 43%. No Hopper fleet-wide MFU band or universal acceptance threshold follows.
As of2024-07
SourceMeta, The Llama 3 Herd of Models, Table 4, July 2024 report. · Table 4: 16,384 GPUs; TP8, CP1, PP16, DP128; sequence 8,192; batch 16/DP; BF16 MFU 41%.
Reviewchecking…review by 2026-07-27 · standard cadence
Recorded changeslast 2026-09-16 · 2 revisions tracked
Claim idbf16-mfu-achieved-pre-training-llama-3-on-16k

Where the guide uses it

← Full numbers register — every date-stamped figure in the guide, with revision history.