15 matches found
EUVD-2026-82534
vLLM through 0.29.0 fails to properly clean up decode-side metadata for rejected inference requests in prefill/decode disaggregated deployments. Remote attackers can submit requests with maxtokens=0 to exhaust decode-worker memory without bound until the worker restarts...
8.7CVSS0.00538EPSS
SaveExploits0References5
20