NVIDIA reports Slinky slurm-operator production clusters exceeding 8,000 GPUs
8,000+ GPUsobserved
| Value kind | observed — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range. |
|---|---|
| Scope | Named NVIDIA production deployment; no performance or readiness guarantee for another installation. |
| As of | 2026-04-09 |
| Source | NVIDIA Developer, “Running Large-Scale GPU Workloads on Kubernetes with Slurm,” 9 April 2026. · Slinky slurm-operator at scale; production clusters run LLM training and multinode inference. |
| Review | checking…review by 2026-08-25 · fast cadence |
| Recorded changes | last 2026-06-29 |
| Claim id | demonstrated-scale-of-slurm-on-kubernetes |
Where the guide uses it
Not quoted in a chapter yet; it is kept in the curated register.
← Full numbers register — every date-stamped figure in the guide, with revision history.