Buffer overflow in Ggml Llama.cpp

CVE-2026-33298

llama.cpp is an inference of several LLM models in C/C++. Prior to b7824, an integer overflow vulnerability in the `ggml_nbytes` function allows an attacker to bypass memory validation by crafting a GGUF file with specific tensor dimensions. This causes `ggml_nbytes` to return a significantly smaller size than required (e.g., 4MB instead of Exabytes), leading to a heap-based buffer overflow when the application subsequently processes the tensor. This vulnerability allows potential Remote Code Execution (RCE) via memory corruption. b7824 contains a fix.

Vulnerability class: Buffer Overflow

Published · last modified .

CVSS v3 metric

CVSS v3 base score 7.8 (High). Vector: CVSS:3.1/AV:L/AC:L/PR:N/UI:R/S:U/C:H/I:H/A:H.

EPSS exploit prediction

EPSS: 0.004 (28.3th percentile), scored .

Very low probability of exploitation in the next 30 days; routine patching cadence is appropriate. 28th percentile — 28.3% of CVEs in the catalogue have a lower EPSS than this one. How to read EPSS.

EPSS trend (30 days)EPSS over the last 30 days for CVE-2026-33298: held from 0.005 to 0.004.

EPSS over last 30 days · oldest: 0.005 · newest: 0.004 · change: -0.001

Affected products

Weakness classification (CWE)

References

Frequently asked questions

What is CVE-2026-33298?
CVE-2026-33298 is a high-severity vulnerability in Ggml Llama.cpp, classified under Heap-based Buffer Overflow. CVSS score: 7.8/10. Published 2026-03-24.
How severe is CVE-2026-33298?
High severity. CVSS v3 base score is 7.8 out of 10.