2 matches found
PT-2026-105921
Summary Current vLLM main lets an inference request choose the PyNvVideoCodec GPU video decoder through media io kwargs.video.video backend, but engine GPU memory reservation is computed only from static startup configuration and VLLM VIDEO LOADER BACKEND. If the server starts with the default...
6.5CVSS5.9AI score0.00583EPSS
SaveExploits1References10
PT-2026-93891
Name of the Vulnerable Software and Affected Versions vLLM versions prior to 0.28.0 Description An issue exists where the engine fails to properly budget GPU memory when a user specifies a GPU decoder at request time. Specifically, request bodies for Chat Completions and Responses can set media i...
6.5CVSS5.8AI score0.00583EPSS
SaveExploits1References13
20