Lucene search
+L

4 matches found

PyPA
PyPA
•added 2026/10/01 4:38 p.m.•13 views

vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation

SummaryCurrent vLLM main lets an inference request choose the PyNvVideoCodec GPU video decoder through mediaiokwargs.video.videobackend, but engine GPU memory reservation is computed only from static startup configuration and VLLMVIDEOLOADERBACKEND. If the server starts with the default...

6.5CVSS5.8AI score0.00548EPSS
SaveExploits1References9Affected Software1
Github Security Blog
Github Security Blog
•added 2026/09/17 5:17 p.m.•18 views

vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation

Summary Current vLLM main lets an inference request choose the PyNvVideoCodec GPU video decoder through mediaiokwargs.video.videobackend, but engine GPU memory reservation is computed only from static startup configuration and VLLMVIDEOLOADERBACKEND. If the server starts with the default...

6.5CVSS5.8AI score0.00548EPSS
SaveExploits1References7Affected Software1
CVE
CVE
•added 2026/09/16 5:49 p.m.•42 views

CVE-2026-69147

vLLM (inference/serving engine for LLMs) prior to 0.28.0 allows a request-selected pynvvideocodec backend to bypass the engine's static VRAM reservation. Specifically, MediaConnector.fetch_video forwards a request-level video_backend override to VideoMediaIO even when startup configuration select...

6.5CVSS5.9AI score0.00548EPSS
SaveExploits1References4Affected Software1
Cvelist
Cvelist
•added 2026/09/16 5:49 p.m.•52 views

CVE-2026-69147 vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation

vLLM is an inference and serving engine for large language models. Prior to 0.28.0, request bodies for Chat Completions and Responses can set mediaiokwargs.video.videobackend to pynvvideocodec, and MediaConnector.fetchvideo forwards that choice to VideoMediaIO even when startup configuration...

6.5CVSS0.00548EPSS
SaveExploits1References4
Rows per page
Query Builder