Lucene search
+L

20 matches found

Vulnrichment
Vulnrichment
•added 2026/10/05 10:46 p.m.•4 views

CVE-2026-105754 vLLM: Scale-out disaggregated multimodal transport trusts caller-supplied features

vLLM is an inference and serving engine for large language models. Prior to 0.30.0, the /inference/v1/generate endpoint in the disaggregated scale-out path accepts caller-supplied tensors in the features.kwargsdata field, cache identifiers in the features.mmhashes field, ranges in the...

6.5CVSS5.9AI score0.00269EPSS
SaveExploits0References4
NVD
NVD
•added 2026/09/17 2:18 p.m.•14 views

CVE-2026-92971

InternLM LMDeploy through 0.17.0 contains a reachable assertion vulnerability in the DistServe decode migration loop that allows unauthenticated attackers to terminate the inference engine. Attackers can submit a migrationrequest with an empty remoteblockids list to trigger an AssertionError that...

8.7CVSS0.00704EPSS
SaveExploits0References6
Cvelist
Cvelist
•added 2026/09/17 1:43 p.m.•37 views

CVE-2026-92971 InternLM LMDeploy through 0.17.0 Assertion Denial of Service

InternLM LMDeploy through 0.17.0 contains a reachable assertion vulnerability in the DistServe decode migration loop that allows unauthenticated attackers to terminate the inference engine. Attackers can submit a migrationrequest with an empty remoteblockids list to trigger an AssertionError that...

8.7CVSS0.00704EPSS
SaveExploits0References6
Packet Storm News
Packet Storm News
•added 2026/09/17 12:00 a.m.•8 views

Inference-Engine Fingerprinting Attacks Are Practical: Exploring Model-Driven Environmental Discovery, Exploitation, and Escape

Frontier AI models are rapidly gaining the ability to exploit vulnerabilities in complex pieces of software. The risk is not theoretical, as evidenced by recent sandbox escapes performed by frontier models at OpenAI and Anthropic. Discussions of how to sandbox inference stack components often foc...

5.9AI score
SaveExploits0
OSV
OSV
•added 2026/08/17 8:17 p.m.•26 views

CVE-2026-73560 vLLM: SSRF + arbitrary local file read in MiMoV2OmniMultiModalProcessor `_fetch_image` and audio loader bypass MediaConnector protections

vLLM is an inference and serving engine for large language models. Prior to 0.26.0, the MiMoV2OmniMultiModalProcessor in vllm/transformersutils/processors/mimov2omni.py passes attacker-controlled image and audio strings through fetchimage, requests.get, and Image.open instead of MediaConnector,...

6.5CVSS5.6AI score
SaveExploits0References6
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/08/13 12:00 a.m.•5 views

CVE-2026-73558: Integer Overflow or Wraparound

vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x 2 d in activationkernels.cu can cause actandmulkernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or...

5.3CVSS6AI score0.00414EPSS
SaveExploits1References7
RedhatCVE
RedhatCVE
•added 2026/08/05 11:12 a.m.•38 views

CVE-2026-44223

A flaw was found in vLLM, an inference and serving engine for large language models LLMs. The extracthiddenstates speculative decoding proposer returns a tensor with an incorrect shape after the first decode step. This can be triggered by a remote attacker sending a request that uses sampling...

7.5CVSS5.1AI score0.00426EPSS
SaveExploits0References5
Microsoft CVE
Microsoft CVE
•added 2026/07/07 8:01 a.m.•17 views

onnx onnxruntime old.cc convPoolShapeInference_opset19 out-of-bounds

...

5.3CVSS5.7AI score0.0043EPSS
SaveExploits0
Positive Technologies
Positive Technologies
•added 2026/04/02 12:00 a.m.•33 views

PT-2026-29877

Name of the Vulnerable Software and Affected Versions vLLM versions 0.5.5 through 0.17.999 Description vLLM, an inference and serving engine for large language models LLMs, exhibits an inconsistency in audio processing. Versions 0.5.5 through 0.17.999 utilize numpy.mean for mono downmixing via...

7.1CVSS5.2AI score0.00476EPSS
SaveExploits0References13
GithubExploit
GithubExploit
•added 2026/01/05 6:58 p.m.•174 views

FoolishScan

Foolish Scan v2.3 Gold Master Context-Aware CTF & Lab Re...

7.1AI score
SaveExploits0
GithubExploit
GithubExploit
•added 2026/01/05 6:58 p.m.•181 views

FoolishScan-

Foolish Scan v2.3 Gold Master Context-Aware CTF & Lab Re...

7.1AI score
SaveExploits0
EUVD
EUVD
•added 2025/10/03 8:07 p.m.•19 views

EUVD-2025-16189

Malicious code in bioql PyPI...

2.6CVSS6.3AI score0.003EPSS
SaveExploits0References4
Mend
Mend
•added 2025/06/24 3:21 a.m.•1 views

CVE-2025-52566

llama.cpp is an inference of several LLM models in C/C++. Prior to version b5721, there is a signed vs. unsigned integer overflow in llama.cpp's tokenizer implementation llamavocab::tokenize src/llama-vocab.cpp:3036 resulting in unintended behavior in tokens copying size comparison. Allowing...

9.3CVSS0.00363EPSS
SaveExploits1References5
NVD
NVD
•added 2025/05/30 7:15 p.m.•45 views

CVE-2025-48942

vLLM is an inference and serving engine for large language models LLMs. In versions 0.8.0 up to but excluding 0.9.0, hitting the /v1/completions API with a invalid jsonschema as a Guided Param kills the vllm server. This vulnerability is similar GHSA-9hcf-v7m4-6m2j/CVE-2025-48943, but for regex...

6.5CVSS0.00551EPSS
SaveExploits1References4
OSV
OSV
•added 2025/05/30 6:38 p.m.•13 views

CVE-2025-48944 vLLM Tool Schema allows DoS via Malformed pattern and type Fields

vLLM is an inference and serving engine for large language models LLMs. In version 0.8.0 up to but excluding 0.9.0, the vLLM backend used with the /v1/chat/completions OpenAPI endpoint fails to validate unexpected or malformed input in the "pattern" and "type" fields when the tools functionality ...

6.5CVSS6.5AI score
SaveExploits0References4
OSV
OSV
•added 2025/05/29 5:15 p.m.•14 views

PYSEC-2025-53

vLLM is an inference and serving engine for large language models LLMs. Prior to version 0.9.0, when a new prompt is processed, if the PageAttention mechanism finds a matching prefix chunk, the prefill process speeds up, which is reflected in the TTFT Time to First Token. These timing differences...

2.6CVSS7AI score0.003EPSS
SaveExploits0References3
CNNVD
CNNVD
•added 2025/03/20 12:00 a.m.•20 views

编号撤回

vLLM is vLLM open source a high throughput and memory efficient inference and service engine for LLM. This CVE number has been withdrawn...

7.6AI score
SaveExploits0References1
CNNVD
CNNVD
•added 2025/02/07 12:00 a.m.•16 views

vLLM 安全漏洞

vLLM is a high throughput and memory efficient inference and service engine for LLM from the vLLM open source. A security vulnerability exists in vLLM that stems from a maliciously constructed statement that could lead to a hash collision, which could lead to cache reuse, which could interfere wi...

2.6CVSS4.3AI score0.00191EPSS
SaveExploits0References3
Snyk
Snyk
•added 2025/02/06 8:00 p.m.•15 views

Use of Weak Hash

Overview vllm is an A high-throughput and memory-efficient inference and serving engine for LLMs Affected versions of this package are vulnerable to Use of Weak Hash due to the use of a predictable constant value in the Python 3.12 built-in hash function. An attacker can interfere with subsequent...

2.6CVSS6.9AI score0.00191EPSS
SaveExploits0References2
Spring Security Advisories
Spring Security Advisories
•added 2024/07/31 12:00 a.m.•48 views

Spring AI with Groq - a blazingly fast AI inference engine

Faster information processing not only informs - it transforms how we perceive and innovate. Spring AI, a powerful framework for integrating AI capabilities into Spring applications, now offers support for Groq - a blazingly fast AI inference engine with support for Tool/Function calling...

6.9AI score
SaveExploits0
Rows per page
Query Builder