Lucene search
+L

32 matches found

CVE
CVE
•added 2026/07/21 8:36 p.m.•40 views

CVE-2026-63764

LMDeploy prior to 0.14.0 contains a server-side request forgery (SSRF) in the _load_http_url function of the media handler. The private-IP guard validates only the original URL; after HTTP redirects, hosts are not re-validated. An unauthenticated attacker can submit a crafted image_url to the cha...

8.6CVSS5.9AI score0.00511EPSS
SaveExploits1References4Affected Software1
Positive Technologies
Positive Technologies
•added 2026/07/21 12:00 a.m.•32 views

PT-2026-62099

Name of the Vulnerable Software and Affected Versions lmdeploy affected versions not specified Description The OpenAI-compatible API server is subject to server-side request forgery, a flaw where the server is tricked into making requests to an unintended location. Unauthenticated attackers can...

8.6CVSS5.7AI score0.00511EPSS
SaveExploits1References13
EUVD
EUVD
•added 2026/07/14 8:09 p.m.•27 views

EUVD-2026-44470

NVIDIA TensorRT-LLM contains a vulnerability in the OpenAI-compatible inference API, where an attacker could cause allocation of GPU resources without limits or throttling. A successful exploit of this vulnerability might lead to denial of service...

6.2CVSS6.1AI score0.0016EPSS
SaveExploits0References2
Cvelist
Cvelist
•added 2026/07/14 8:09 p.m.•57 views

CVE-2026-24271

NVIDIA TensorRT-LLM contains a vulnerability in the OpenAI-compatible inference API, where an attacker could cause allocation of GPU resources without limits or throttling. A successful exploit of this vulnerability might lead to denial of service...

6.2CVSS0.0016EPSS
SaveExploits0References2
attackerkb
attackerkb
•added 2026/07/14 8:08 p.m.•19 views

CVE-2026-47475

NVIDIA TensorRT-LLM contains a vulnerability in the OpenAI-compatible inference API where an attacker could trigger a reachable assertion in the sampler thread. A successful exploit of this vulnerability might lead to denial of service...

6.2CVSS6.1AI score0.0016EPSS
SaveExploits0References3
CVE
CVE
•added 2026/07/14 8:08 p.m.•31 views

CVE-2026-47475

Summary of CVE-2026-47475 (NVIDIA TensorRT-LLM) : A vulnerability in the OpenAI-compatible inference API could allow an attacker to trigger a reachable assertion in the sampler thread, potentially causing a denial of service. The issue is scoped to NVIDIA TensorRT-LLM, with local attack vector an...

6.2CVSS5.8AI score0.0016EPSS
SaveExploits0References2
PyPA
PyPA
•added 2026/06/11 10:16 a.m.•18 views

PYSEC-2026-2302

vLLM versions 0.8.0 and later are vulnerable to an Out-of-Memory OOM Denial of Service DoS attack due to unbounded frame count processing in the VideoMediaIO.loadbase64 method. When processing video/jpeg data URLs, the method splits the base64 data string on commas to extract individual JPEG fram...

7.5CVSS7.1AI score0.00896EPSS
SaveExploits1References6Affected Software1
CVE
CVE
•added 2026/06/11 8:31 a.m.•113 views

CVE-2026-5497

CVE-2026-5497 affects vLLM 0.8.0 and later, where VideoMediaIO.load_base64() can perform unbounded frame processing for video/jpeg data URLs, leading to an Out-of-Memory DoS. An attacker can craft a single API request with thousands of comma-separated base64 JPEG frames, causing the server to dec...

7.5CVSS7.1AI score0.00896EPSS
SaveExploits1References5Affected Software1
EUVD
EUVD
•added 2026/04/06 3:40 p.m.•25 views

EUVD-2026-19351

vLLM is an inference and serving engine for large language models LLMs. From 0.1.0 to before 0.19.0, a Denial of Service vulnerability exists in the vLLM OpenAI-compatible API server. Due to the lack of an upper bound validation on the n parameter in the ChatCompletionRequest and CompletionReques...

6.5CVSS5.9AI score0.00766EPSS
SaveExploits0References3
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/04/06 12:00 a.m.•1 views

CVE-2026-34756: Allocation of Resources Without Limits or Throttling

vLLM is an inference and serving engine for large language models LLMs. From 0.1.0 to before 0.19.0, a Denial of Service vulnerability exists in the vLLM OpenAI-compatible API server. Due to the lack of an upper bound validation on the n parameter in the ChatCompletionRequest and CompletionReques...

6.5CVSS6AI score0.00766EPSS
SaveExploits0References6
Github Security Blog
Github Security Blog
•added 2025/04/15 9:21 p.m.•208 views

vLLM vulnerable to Denial of Service by abusing xgrammar cache

Impact This report is to highlight a vulnerability in XGrammar, a library used by the structured output feature in vLLM. The XGrammar advisory is here: https://github.com/mlc-ai/xgrammar/security/advisories/GHSA-389x-67px-mjg3 The xgrammar library is the default backend used by vLLM to support...

6.8AI score
SaveExploits0References5Affected Software1
OSV
OSV
•added 2025/03/19 4:15 p.m.•17 views

PYSEC-2025-223

vLLM is a high-throughput and memory-efficient inference and serving engine for LLMs. The outlines library is one of the backends used by vLLM to support structured output a.k.a. guided decoding. Outlines provides an optional cache for its compiled grammars on the local filesystem. This cache has...

6.5CVSS6.6AI score0.00461EPSS
SaveExploits0References3
Rows per page
Query Builder