Lucene search
+L

3 matches found

PyPA
PyPA
added 2025/05/29 5:15 p.m.18 views

PYSEC-2025-53

vLLM is an inference and serving engine for large language models LLMs. Prior to version 0.9.0, when a new prompt is processed, if the PageAttention mechanism finds a matching prefix chunk, the prefill process speeds up, which is reflected in the TTFT Time to First Token. These timing differences...

2.6CVSS6.8AI score0.00257EPSS
SaveExploits0References4Affected Software1
OSV
OSV
added 2025/05/29 5:15 p.m.8 views

PYSEC-2025-53

vLLM is an inference and serving engine for large language models LLMs. Prior to version 0.9.0, when a new prompt is processed, if the PageAttention mechanism finds a matching prefix chunk, the prefill process speeds up, which is reflected in the TTFT Time to First Token. These timing differences...

2.6CVSS7AI score0.00257EPSS
SaveExploits0References3
CVE
CVE
added 2025/05/29 4:32 p.m.194 views

CVE-2025-46570

The CVE-2025-46570 entry concerns vLLM (inference/serving engine). The concrete detail across connected records shows a vulnerability in the PageAttention-based prefill path: when a new prompt is processed, a matching prefix chunk can accelerate prefill, creating timing differences (TTFT) that co...

2.6CVSS3.6AI score0.00257EPSS
SaveExploits0References3Affected Software1
Rows per page
Query Builder