GPU-cloud SLA example: node uptime / rack uptime, with penalties (ClusterMAX 2.0 example figures; SemiAnalysis sets no baseline)
99% / 95%observed
| Value kind | observed — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range. |
|---|---|
| Scope | Illustrative provider contract structure reported by ClusterMAX 2.0; not a market baseline or recommendation. |
| Caveat | Uptime definition, exclusions, measurement window, spare entitlement, service credits, and penalties remain contract-specific. |
| As of | 2025 |
| Source | SemiAnalysis ClusterMAX · SLA examples: one provider approach commits 99% node uptime and 95% rack uptime (rack defined as 16 of 18 nodes, or 64 of 72 GPUs); another commits 99% node uptime only and allocates 16 of 18 nodes. |
| Review | checking…review by 2026-10-26 · standard cadence |
| Recorded changes | last 2026-07-21 · 2 revisions tracked |
| Claim id | gpu-cloud-sla-baseline-node-uptime-rack-uptime |
Where the guide uses it
← Full numbers register — every date-stamped figure in the guide, with revision history.