Lucene search
+L

661 matches found

RedhatCVE
RedhatCVE
added 2026/06/29 4:39 a.m.13 views

CVE-2026-53923

A flaw was found in vLLM. Integer truncation of tensor dimensions in vLLM's GGUF dequantize kernels leads to partial tensor processing. This results in the output tensor retaining previously used GPU memory, which, in multi-tenant inference deployments, can expose sensitive tensor data from other...

7.5CVSS5.7AI score0.00281EPSS
SaveExploits0References6
RedhatCVE
RedhatCVE
added 2026/06/26 7:34 a.m.15 views

CVE-2026-54236

A flaw was found in vLLM, an inference and serving engine for large language models LLMs. An unauthenticated attacker can exploit this vulnerability by sending specially crafted malformed image bytes through the Anthropic Messages API. This action causes an error message to be generated that...

5.3CVSS5.6AI score0.00823EPSS
SaveExploits1References6
RedhatCVE
RedhatCVE
added 2026/06/26 7:9 a.m.10 views

CVE-2026-54232

A flaw was found in vLLM, an inference and serving engine for large language models LLMs. This vulnerability, a dependency confusion attack, allows a remote attacker to execute arbitrary code with root privileges during the Docker build process. By exploiting this, an attacker can compromise the...

8.8CVSS6.1AI score0.00563EPSS
SaveExploits1References4
RedhatCVE
RedhatCVE
added 2026/06/24 3:19 p.m.12 views

CVE-2026-48746

A flaw was found in vLLM, an inference and serving engine for large language models LLMs. This vulnerability, residing in ASGI web servers and Starlette's trust in them, allows an attacker to bypass the OpenAI API Authentication Middleware. This bypass enables unauthorized access to the API witho...

9.1CVSS5.8AI score0.01152EPSS
SaveExploits0References6
Chainguard
Chainguard
added 2026/06/23 8:21 p.m.8 views

CVE-2026-48746 vulnerabilities

Vulnerabilities for packages: py3-vllm-cuda-12.9, py3-vllm-cuda-12.4, tritonserver-backend-vllm-cuda-13.0, py3-vllm-cuda-13.0...

9.1CVSS6.9AI score0.01152EPSS
SaveExploits0
Chainguard
Chainguard
added 2026/06/23 8:21 p.m.13 views

GHSA-94F4-HR76-P5J6 vulnerabilities

Vulnerabilities for packages: py3-vllm-cuda-12.9, py3-vllm-cuda-12.4, tritonserver-backend-vllm-cuda-13.0, py3-vllm-cuda-13.0...

5.8AI score
SaveExploits0
Chainguard
Chainguard
added 2026/06/23 8:16 a.m.17 views

GHSA-4XGF-CPJX-PC3J vulnerabilities

Vulnerabilities for packages: tritonserver-backend-vllm-cuda-12.9, prefect, vllm-cuda-13.2, vllm-openai-cuda-13.0, litellm, azureml-inference-server-http, azureml-inference-server-http-fips, airflow, mcp-atlassian, lmcache-cuda-12.8, open-webui, airflow-core, airflow-postgres-fips,...

5.8AI score
SaveExploits0
NVD
NVD
added 2026/06/22 11:16 p.m.11 views

CVE-2026-41523

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, an assert-based security check in vLLM's activation function loading allows any unauthenticated attacker to achieve arbitrary code execution on the server by publishing a malicious HuggingFace model, when vLL...

7.5CVSS0.00746EPSS
SaveExploits1References8
NVD
NVD
added 2026/06/22 11:16 p.m.12 views

CVE-2026-47155

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, vLLM's revision pinning controls do not consistently apply to all artifacts loaded for a model. A deployment that supplies --revision or --code-revision can still load dynamic code, GGUF files, image...

6.5CVSS0.0021EPSS
SaveExploits0References4
OSV
OSV
added 2026/06/22 11:16 p.m.3 views

PYSEC-2026-2301

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, vLLM's revision pinning controls do not consistently apply to all artifacts loaded for a model. A deployment that supplies --revision or --code-revision can still load dynamic code, GGUF files, image...

6.5CVSS6.1AI score0.0021EPSS
SaveExploits0References4
NVD
NVD
added 2026/06/22 11:16 p.m.12 views

CVE-2026-48746

vLLM is an inference and serving engine for large language models LLMs. From 0.3.0 until 0.22.0, a vulnerability in ASGI web servers and starlette's trust on those web servers enables an authentication bypass of the OpenAI API AuthenticationMiddleware. It allows to use the API without providing t...

9.1CVSS0.01152EPSS
SaveExploits0References14
CVE
CVE
added 2026/06/22 10:18 p.m.52 views

CVE-2026-41523

vLLM prior to 0.22.0 is affected by an assert-based security check in the activation function loading that can permit arbitrary code execution when a malicious HuggingFace model is loaded and vLLM runs in Python optimized mode. The attacker-controlled inputs are the activation function names from...

7.5CVSS6.5AI score0.00746EPSS
SaveExploits1References8Affected Software1
ATTACKERKB
ATTACKERKB
added 2026/06/22 10:18 p.m.9 views

CVE-2026-41523

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, an assert-based security check in vLLM's activation function loading allows any unauthenticated attacker to achieve arbitrary code execution on the server by publishing a malicious HuggingFace model, when vLL...

7.5CVSS6.5AI score0.00746EPSS
SaveExploits1References4Affected Software1
Vulnrichment
Vulnrichment
added 2026/06/22 10:18 p.m.10 views

CVE-2026-41523 vLLM: Security Check Bypass via assert Statement in Activation Function Loading Allows Arbitrary Code Execution

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, an assert-based security check in vLLM's activation function loading allows any unauthenticated attacker to achieve arbitrary code execution on the server by publishing a malicious HuggingFace model, when vLL...

7.5CVSS6.5AI score0.00746EPSS
SaveExploits1References3
OSV
OSV
added 2026/06/22 10:18 p.m.5 views

CVE-2026-41523 vLLM: Security Check Bypass via assert Statement in Activation Function Loading Allows Arbitrary Code Execution

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.0, an assert-based security check in vLLM's activation function loading allows any unauthenticated attacker to achieve arbitrary code execution on the server by publishing a malicious HuggingFace model, when vLL...

7.5CVSS6.6AI score0.00746EPSS
SaveExploits1References10
Cvelist
Cvelist
added 2026/06/22 10:16 p.m.29 views

CVE-2026-54232 vLLM: Dependency Confusion Vulnerability in vLLM Dockerfile

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.1, the vLLM Dockerfile is vulnerable to a dependency confusion attack through the flashinfer-jit-cache package. The package is installed from a custom index flashinfer.ai/whl/ using --extra-index-url, but the...

8.8CVSS0.00563EPSS
SaveExploits1References1
ATTACKERKB
ATTACKERKB
added 2026/06/22 10:16 p.m.8 views

CVE-2026-54232

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.1, the vLLM Dockerfile is vulnerable to a dependency confusion attack through the flashinfer-jit-cache package. The package is installed from a custom index flashinfer.ai/whl/ using --extra-index-url, but the...

8.8CVSS6.2AI score0.00563EPSS
SaveExploits1References2Affected Software1
OSV
OSV
added 2026/06/22 10:16 p.m.6 views

CVE-2026-54232 vLLM: Dependency Confusion Vulnerability in vLLM Dockerfile

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.22.1, the vLLM Dockerfile is vulnerable to a dependency confusion attack through the flashinfer-jit-cache package. The package is installed from a custom index flashinfer.ai/whl/ using --extra-index-url, but the...

8.8CVSS6.3AI score0.00563EPSS
SaveExploits1References3
Cvelist
Cvelist
added 2026/06/22 10:10 p.m.41 views

CVE-2026-54233 vLLM: OOM Denial of Service via Audio Decompression Bomb

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.23.1rc0, vLLM's /v1/audio/transcriptions endpoint limits compressed upload size but not decoded PCM output. A 25MB OPUS file expands to 14.9GB of float32 PCM at decode time. This vulnerability is fixed in 0.23.1rc0...

6.5CVSS0.00422EPSS
SaveExploits0References2
OSV
OSV
added 2026/06/22 10:10 p.m.4 views

CVE-2026-54233 vLLM: OOM Denial of Service via Audio Decompression Bomb

vLLM is an inference and serving engine for large language models LLMs. Prior to 0.23.1rc0, vLLM's /v1/audio/transcriptions endpoint limits compressed upload size but not decoded PCM output. A 25MB OPUS file expands to 14.9GB of float32 PCM at decode time. This vulnerability is fixed in 0.23.1rc0...

6.5CVSS5.9AI score0.00422EPSS
SaveExploits0References4
Rows per page
Query Builder