The Definitive Guide toAI Data Centers
Ask the GuideAboutAccount
Guide › Numbers register › Claim

LLM-response data recoverable per query via LeftoverLocals (CVE-2023-4969) from un-scrubbed GPU local memory

≈181 MBobserved

Value kindobserved — Reported measurements, counts, and specifications keep the precision and scope stated by their source; an exact specification is not treated as a range.
As of2024
SourceTrail of Bits · Introductory findings: about 5.5 MB per GPU invocation and about 181 MB for each llama.cpp 7B-model query
Reviewchecking…review by 2026-10-07 · standard cadence
Recorded changeslast 2026-06-29
Claim idllm-response-data-recoverable-per-query-via

Where the guide uses it

← Full numbers register — every date-stamped figure in the guide, with revision history.