Information disclosure in Vllm
CVE-2026-53923
vLLM is an inference and serving engine for large language models (LLMs). From 0.5.5 until 0.23.1rc0, integer truncation of tensor dimensions in vLLM's GGUF dequantize kernels (csrc/quantization/gguf/gguf_kernel.cu) causes partial tensor processing. The output tensor is allocated at full size via torch::empty (uninitialized memory), but the dequantize CUDA kernel processes only a truncated number of elements. The unfilled portion of the output tensor retains whatever was previously in GPU memory. In multi-tenant inference deployments, this residual GPU memory may contain tensor data from other users' inference requests, constituting information disclosure. This vulnerability is fixed in 0.23.1rc0.
Vulnerability class: Information Disclosure
Published · last modified .
CVSS v3 metric
CVSS v3 base score 7.5 (High). Vector: CVSS:3.1/AV:N/AC:L/PR:N/UI:N/S:U/C:H/I:N/A:N.
EPSS exploit prediction
EPSS: 0.005 (39.3th percentile), scored .
Very low probability of exploitation in the next 30 days; routine patching cadence is appropriate. 39th percentile — 39.3% of CVEs in the catalogue have a lower EPSS than this one. How to read EPSS.
EPSS over last 30 days · oldest: 0.005 · newest: 0.005 · change: 0.000
Affected products
- Vllm
- Vllm-Project Vllm — versions >= 0.5.5, < 0.23.1rc0
Weakness classification (CWE)
References
- github.com/vllm-project/vllm/security/advisories/GHSA-5jv2-g5wq-cmr4 (Third Party Advisory)
- github.com/vllm-project/vllm/pull/44971 (Issue Tracking)
- github.com/vllm-project/vllm/commit/f219788f91952827132fa4fdf916427cd20d225e (Patch)
Frequently asked questions
- What is CVE-2026-53923?
- CVE-2026-53923 is a high-severity vulnerability in Vllm, classified under Information Disclosure. CVSS score: 7.5/10. Published 2026-06-22.
- How severe is CVE-2026-53923?
- High severity. CVSS v3 base score is 7.5 out of 10.