98 matches found
EUVD-2026-58066
vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...
5.3CVSS5.4AI score
SaveExploits0References5
20