2 matches found
CVE-2026-44223: Incorrect Calculation of Buffer Size
vLLM is an inference and serving engine for large language models LLMs. From 0.18.0 to before 0.20.0, the extracthiddenstates speculative decoding proposer in vLLM returns a tensor with an incorrect shape after the first decode step, causing a RuntimeError that crashes the EngineCore process. The...
6.5CVSS5.9AI score0.00426EPSS
SaveExploits0References7
vLLM 安全漏洞
vLLM is an open-source LLM-based inference and service engine that features high throughput and efficient memory usage. Versions of vLLM prior to 0.20.0 contained a security vulnerability. This vulnerability stemmed from the extracthiddenstates speculative decoding proposal, which returned tensor...
6.5CVSS5.8AI score0.00426EPSS
SaveExploits0References1
20