Lucene search
+L

6 matches found

NVD
NVD
•added 2026/08/13 3:20 p.m.•13 views

CVE-2026-73558

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...

5.3CVSS0.00414EPSS
SaveExploits1References5
attackerkb
attackerkb
•added 2026/08/13 3:06 p.m.•14 views

CVE-2026-73558

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...

5.3CVSS5.3AI score0.00414EPSS
SaveExploits1References6Affected Software1
Cvelist
Cvelist
•added 2026/08/13 3:06 p.m.•42 views

CVE-2026-73558 vLLM: Cross-User Data Leak Vulnerability

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...

5.3CVSS0.00414EPSS
SaveExploits1References5
OSV
OSV
•added 2026/08/13 3:06 p.m.•49 views

CVE-2026-73558 vLLM: Cross-User Data Leak Vulnerability

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...

5.3CVSS5.5AI score
SaveExploits0References7
Positive Technologies
Positive Technologies
•added 2026/08/13 12:00 a.m.•13 views

PT-2026-71673

Name of the Vulnerable Software and Affected Versions vLLM versions prior to 0.27.0 Description An integer overflow occurs in the calculation blockIdx.x 2 d within the activation kernels.cu file. This flaw can cause the act and mul kernel function to consume input from another user within the sam...

5.3CVSS5.7AI score0.00414EPSS
SaveExploits1References15
Github Security Blog
Github Security Blog
•added 2026/06/17 2:03 p.m.•45 views

vLLM: GGUF dequantize kernel int truncation exposes uninitialized GPU memory in multi-tenant serving

Summary Integer truncation of tensor dimensions in vLLM's GGUF dequantize kernels csrc/quantization/gguf/ggufkernel.cu causes partial tensor processing. The output tensor is allocated at full size via torch::empty uninitialized memory, but the dequantize CUDA kernel processes only a truncated...

7.5CVSS5.6AI score0.00484EPSS
SaveExploits0References8Affected Software1
Rows per page
Query Builder