Lucene search
+L

171 matches found

OSV
OSV
added 2026/07/13 3:46 p.m.7 views

PYSEC-2026-3403 vLLM: GGUF dequantize kernel int truncation exposes uninitialized GPU memory in multi-tenant serving

Summary Integer truncation of tensor dimensions in vLLM's GGUF dequantize kernels csrc/quantization/gguf/ggufkernel.cu causes partial tensor processing. The output tensor is allocated at full size via torch::empty uninitialized memory, but the dequantize CUDA kernel processes only a truncated...

5.3CVSS6.2AI score0.00281EPSS
SaveExploits0References7
Rows per page
Query Builder