1 matches found
vLLM: Derender endpoints decode caller-supplied GenerateResponse token IDs without output bounds
SummaryThe /v1/completions/derender and /v1/chat/completions/derender endpoints accept caller-supplied GenerateResponse objects and postprocess every nested choices.tokenids list directly. Unlike the normal render/generate path, derender does not enforce model context length, resolved maxtokens,...
4.3CVSS5.8AI score0.00374EPSS
SaveExploits1References8Affected Software1
20