Lucene search
+L

661 matches found

Chainguard
Chainguard
added 2025/12/01 7:44 p.m.7 views

GHSA-Q279-JHRF-CC6V vulnerabilities

Vulnerabilities for packages: py3-vllm-cuda-12.4, tritonserver-backend-vllm-cuda-12.9, airflow...

5.8AI score
SaveExploits0
Chainguard
Chainguard
added 2025/12/01 7:44 p.m.4 views

GHSA-MRW7-HF4F-83PF vulnerabilities

Vulnerabilities for packages: tritonserver-backend-vllm-cuda-12.9...

7AI score
SaveExploits0
Chainguard
Chainguard
added 2025/12/01 7:44 p.m.24 views

CVE-2025-62426 vulnerabilities

Vulnerabilities for packages: tritonserver-backend-vllm-cuda-12.9...

6.5CVSS6.7AI score0.00356EPSS
SaveExploits0
Veracode
Veracode
added 2025/11/24 3:37 p.m.7 views

Server-Side Request Forgery (SSRF)

vllm is vulnerable to Server-Side Request Forgery SSRF. The vulnerability is due to insufficient restrictions on user-supplied URLs in the MediaConnector class’s loadfromurl and loadfromurlasync methods, which allows an attacker to coerce the server into making arbitrary internal network requests...

7.1CVSS7.2AI score0.00229EPSS
SaveExploits0References6Affected Software1
NVD
NVD
added 2025/11/21 2:15 a.m.8 views

CVE-2025-62372

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, users can crash the vLLM engine serving multimodal models by passing multimodal embedding inputs with correct ndim but incorrect shape e.g. hidden dimension is wrong, regardless of whether...

8.3CVSS0.0037EPSS
SaveExploits0References4
NVD
NVD
added 2025/11/21 2:15 a.m.11 views

CVE-2025-62164

vLLM is an inference and serving engine for large language models LLMs. From versions 0.10.2 to before 0.11.1, a memory corruption vulnerability could lead to a crash denial-of-service and potentially remote code execution RCE, exists in the Completions API endpoint. When processing user-supplied...

8.8CVSS0.0093EPSS
SaveExploits0References3
Cvelist
Cvelist
added 2025/11/21 1:22 a.m.16 views

CVE-2025-62372 vLLM vulnerable to DoS with incorrect shape of multimodal embedding inputs

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, users can crash the vLLM engine serving multimodal models by passing multimodal embedding inputs with correct ndim but incorrect shape e.g. hidden dimension is wrong, regardless of whether...

8.3CVSS0.0037EPSS
SaveExploits0References4
Vulnrichment
Vulnrichment
added 2025/11/21 1:22 a.m.3 views

CVE-2025-62372 vLLM vulnerable to DoS with incorrect shape of multimodal embedding inputs

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, users can crash the vLLM engine serving multimodal models by passing multimodal embedding inputs with correct ndim but incorrect shape e.g. hidden dimension is wrong, regardless of whether...

8.3CVSS6.5AI score0.0037EPSS
SaveExploits0References4
Vulnrichment
Vulnrichment
added 2025/11/21 1:21 a.m.9 views

CVE-2025-62426 vLLM vulnerable to DoS via large Chat Completion or Tokenization requests with specially crafted `chat_template_kwargs`

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, the /v1/chat/completions and /tokenize endpoints allow a chattemplatekwargs request parameter that is used in the code before it is properly validated against the chat template. With the...

6.5CVSS6.8AI score0.00356EPSS
SaveExploits0References5
Cvelist
Cvelist
added 2025/11/21 1:21 a.m.19 views

CVE-2025-62426 vLLM vulnerable to DoS via large Chat Completion or Tokenization requests with specially crafted `chat_template_kwargs`

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, the /v1/chat/completions and /tokenize endpoints allow a chattemplatekwargs request parameter that is used in the code before it is properly validated against the chat template. With the...

6.5CVSS0.00356EPSS
SaveExploits0References5
OSV
OSV
added 2025/11/21 1:21 a.m.7 views

CVE-2025-62426 vLLM vulnerable to DoS via large Chat Completion or Tokenization requests with specially crafted `chat_template_kwargs`

vLLM is an inference and serving engine for large language models LLMs. From version 0.5.5 to before 0.11.1, the /v1/chat/completions and /tokenize endpoints allow a chattemplatekwargs request parameter that is used in the code before it is properly validated against the chat template. With the...

6.5CVSS7AI score0.00356EPSS
SaveExploits0References7
CVE
CVE
added 2025/11/21 1:18 a.m.51 views

CVE-2025-62164

The CVE affects vLLM (inference/serving engine) before 0.11.1, where the Completions API loads user-supplied prompt embeddings with torch.load() lacking proper validation. A PyTorch 2.8.0 change disables sparse-tensor invariants checks, allowing crafted tensors to bypass bounds checks and trigger...

8.8CVSS7.8AI score0.0093EPSS
SaveExploits0References3Affected Software1
Vulnrichment
Vulnrichment
added 2025/11/21 1:18 a.m.3 views

CVE-2025-62164 VLLM deserialization vulnerability leading to DoS and potential RCE

vLLM is an inference and serving engine for large language models LLMs. From versions 0.10.2 to before 0.11.1, a memory corruption vulnerability could lead to a crash denial-of-service and potentially remote code execution RCE, exists in the Completions API endpoint. When processing user-supplied...

8.8CVSS7.8AI score0.0093EPSS
SaveExploits0References3
Cvelist
Cvelist
added 2025/11/21 1:18 a.m.18 views

CVE-2025-62164 VLLM deserialization vulnerability leading to DoS and potential RCE

vLLM is an inference and serving engine for large language models LLMs. From versions 0.10.2 to before 0.11.1, a memory corruption vulnerability could lead to a crash denial-of-service and potentially remote code execution RCE, exists in the Completions API endpoint. When processing user-supplied...

8.8CVSS0.0093EPSS
SaveExploits0References3
OSV
OSV
added 2025/11/21 1:18 a.m.5 views

CVE-2025-62164 VLLM deserialization vulnerability leading to DoS and potential RCE

vLLM is an inference and serving engine for large language models LLMs. From versions 0.10.2 to before 0.11.1, a memory corruption vulnerability could lead to a crash denial-of-service and potentially remote code execution RCE, exists in the Completions API endpoint. When processing user-supplied...

8.8CVSS8.1AI score0.0093EPSS
SaveExploits0References5
CNNVD
CNNVD
added 2025/11/21 12:0 a.m.7 views

vLLM 缓冲区错误漏洞

vLLM is a vLLM open source high throughput and memory efficient inference and service engine for LLM. A buffer error vulnerability exists in vLLM versions 0.10.2 through prior to 0.11.1, which stems from the presence of a memory corruption in the Completions API endpoint that could lead to a cras...

8.8CVSS7.9AI score0.0093EPSS
SaveExploits0References3
Github Security Blog
Github Security Blog
added 2025/11/20 9:26 p.m.11 views

vLLM vulnerable to DoS via large Chat Completion or Tokenization requests with specially crafted `chat_template_kwargs`

Summary The /v1/chat/completions and /tokenize endpoints allow a chattemplatekwargs request parameter that is used in the code before it is properly validated against the chat template. With the right chattemplatekwargs parameters, it is possible to block processing of the API server for long...

6.5CVSS6.8AI score0.00356EPSS
SaveExploits0References10Affected Software1
Snyk
Snyk
added 2025/11/20 9:26 p.m.11 views

Allocation of Resources Without Limits or Throttling

Overview vllm is an A high-throughput and memory-efficient inference and serving engine for LLMs Affected versions of this package are vulnerable to Allocation of Resources Without Limits or Throttling via the applyhfchattemplate method. An authenticated user can cause the server to become...

7.1CVSS6.9AI score0.00356EPSS
SaveExploits0References2
OSV
OSV
added 2025/11/20 9:26 p.m.3 views

GHSA-69J4-GRXJ-J64P vLLM vulnerable to DoS via large Chat Completion or Tokenization requests with specially crafted `chat_template_kwargs`

Summary The /v1/chat/completions and /tokenize endpoints allow a chattemplatekwargs request parameter that is used in the code before it is properly validated against the chat template. With the right chattemplatekwargs parameters, it is possible to block processing of the API server for long...

6.5CVSS6.1AI score0.00356EPSS
SaveExploits0References10
Github Security Blog
Github Security Blog
added 2025/11/20 8:59 p.m.15 views

vLLM deserialization vulnerability leading to DoS and potential RCE

Summary A memory corruption vulnerability that leading to a crash denial-of-service and potentially remote code execution RCE exists in vLLM versions 0.10.2 and later, in the Completions API endpoint. When processing user-supplied prompt embeddings, the endpoint loads serialized tensors using...

8.8CVSS8.3AI score0.0093EPSS
SaveExploits0References8Affected Software1
Rows per page
Query Builder