137 matches found
vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation
SummaryCurrent vLLM main lets an inference request choose the PyNvVideoCodec GPU video decoder through mediaiokwargs.video.videobackend, but engine GPU memory reservation is computed only from static startup configuration and VLLMVIDEOLOADERBACKEND. If the server starts with the default...
6.5CVSS5.8AI score0.00548EPSS
SaveExploits0References9Affected Software1
20