Lucene search
+L

48 matches found

Kitploit
Kitploit
•added 2026/10/07 3:48 p.m.•11 views

CacheTrap

CacheTrap Repository for "CacheTrap: Unveiling a Stealthier Gray-Box Trojan against LLMs" , IEEE/ACM International Conference on Computer-Aided Design ICCAD, 2026. arXiv CacheTrap searches for a single bit flip in the KV cache of a fine-tuned LLM classifier and measures the resulting attack succe...

6.3AI score
SaveExploits0
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/10/02 12:00 a.m.•5 views

CVE-2026-103765: Missing Authentication for Critical Function

Mooncake through 0.3.13.post1 contains a missing authentication vulnerability in the HTTP metadata server /metadata handler that allows unauthenticated attackers to read, overwrite, and delete transfer engine metadata keys. Attackers can poison segment descriptors such as tcpdataport or re-create...

9.4CVSS5.9AI score0.005EPSS
SaveExploits0References2
OSV
OSV
•added 2026/10/01 11:19 p.m.•5 views

CVE-2026-103765 Mooncake through 0.3.13.post1 Missing Authentication in HTTP Metadata Server

Mooncake through 0.3.13.post1 contains a missing authentication vulnerability in the HTTP metadata server /metadata handler that allows unauthenticated attackers to read, overwrite, and delete transfer engine metadata keys. Attackers can poison segment descriptors such as tcpdataport or re-create...

8.8CVSS5.9AI score
SaveExploits0References6
OSV
OSV
•added 2026/10/01 11:19 p.m.•7 views

CVE-2026-103764 Mooncake transfer engine before 0.3.13 Unauthenticated Arbitrary Memory Read/Write via TCP Transport

Mooncake transfer engine before 0.3.13 contains an untrusted pointer dereference in ServerSession::readHeader that allows unauthenticated attackers to read and write arbitrary process memory via the TCP transport data port. Attackers can send a crafted SessionHeader with arbitrary addr and size...

9.3CVSS6.2AI score
SaveExploits0References8
Positive Technologies
Positive Technologies
•added 2026/10/01 12:00 a.m.•13 views

PT-2026-104057

Mooncake transfer engine before 0.3.13 contains an untrusted pointer dereference in ServerSession::readHeader that allows unauthenticated attackers to read and write arbitrary process memory via the TCP transport data port. Attackers can send a crafted SessionHeader with arbitrary addr and size...

9.8CVSS6.1AI score0.00636EPSS
SaveExploits0References7
SUSE CVE
SUSE CVE
•added 2026/09/23 12:13 a.m.•9 views

SUSE CVE-2026-94627

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS5.9AI score0.00627EPSS
SaveExploits0References3
RedhatCVE
RedhatCVE
•added 2026/09/22 4:44 p.m.•15 views

CVE-2026-94627

A flaw was found in vLLM. A remote attacker can exploit a vulnerability in the Mooncake connector's management of GPU Key-Value KV cache block ownership. By submitting completion requests with multiple prompts that share a single transfer ID, attackers can cause orphaned KV cache blocks to...

8.7CVSS5.9AI score0.00627EPSS
SaveExploits0References7
EUVD
EUVD
•added 2026/09/22 12:30 a.m.•14 views

EUVD-2026-84221

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS5.9AI score0.00627EPSS
SaveExploits0References5
Snyk
Snyk
•added 2026/09/21 11:22 p.m.•25 views

Improper Handling of Exceptional Conditions

Overview vllm is an A high-throughput and memory-efficient inference and serving engine for LLMs Affected versions of this package are vulnerable to Improper Handling of Exceptional Conditions in mooncakeconnector.py via the Mooncake KV transfer connector, when a remote KV cache load fails during...

6.9CVSS5.9AI score0.00521EPSS
SaveExploits0References2
OSV
OSV
•added 2026/09/21 10:04 p.m.•16 views

CVE-2026-94627 vLLM through 0.29.0 GPU KV Cache Leak via Mooncake Transfer ID Collision

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS5.8AI score
SaveExploits0References6
Cvelist
Cvelist
•added 2026/09/21 10:04 p.m.•51 views

CVE-2026-94627 vLLM through 0.29.0 GPU KV Cache Leak via Mooncake Transfer ID Collision

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS0.00627EPSS
SaveExploits0References4
Vulnrichment
Vulnrichment
•added 2026/09/21 10:04 p.m.•16 views

CVE-2026-94627 vLLM through 0.29.0 GPU KV Cache Leak via Mooncake Transfer ID Collision

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS5.9AI score0.00627EPSS
SaveExploits0References4
Positive Technologies
Positive Technologies
•added 2026/09/21 12:00 a.m.•11 views

PT-2026-96317

Name of the Vulnerable Software and Affected Versions vLLM versions prior to 0.29.1 Description The Mooncake connector fails to properly manage GPU KV cache block ownership during prefill/decode disaggregated deployments when concurrent child requests share a single transfer ID. An attacker can...

8.7CVSS5.8AI score0.00627EPSS
SaveExploits0References10
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/09/21 12:00 a.m.•9 views

CVE-2026-94627: Missing Release of Memory after Effective Lifetime

vLLM Mooncake connector through 0.29.0 fails to properly manage GPU KV cache block ownership when concurrent child requests share a single transfer ID in prefill/decode disaggregated deployments. Attackers can trigger GPU memory exhaustion by submitting completion requests with multiple prompts,...

8.7CVSS5.9AI score0.00627EPSS
SaveExploits0References4
OSV
OSV
•added 2026/09/16 5:49 p.m.•11 views

CVE-2026-69147 vLLM: Request-selected PyNvVideoCodec GPU decode bypasses static VRAM reservation

vLLM is an inference and serving engine for large language models. Prior to 0.28.0, request bodies for Chat Completions and Responses can set mediaiokwargs.video.videobackend to pynvvideocodec, and MediaConnector.fetchvideo forwards that choice to VideoMediaIO even when startup configuration...

6.5CVSS5.8AI score
SaveExploits0References6
Packet Storm News
Packet Storm News
•added 2026/09/06 12:00 a.m.•9 views

Characterizing Contention-Induced Reliability Collapse in KV-Cache Timing Side Channels for Multi-Tenant LLM Serving

Shared key--value KV cache reuse improves large language model LLM serving, but it can also create a timing side channel that reveals whether a prefix is already cached. Previous work shows that such attacks are possible, but their reliability under realistic multi-tenant contention is less...

5.8AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/09/03 12:52 a.m.•17 views

llama-cpp-security-patches — Updated!

llama.cpp Security Patches Security patches for unpatched vulnerabilities in llama.cpp, discovered by Cyera Research. Background Between July 2025 and June 2026, we reported 10 vulnerabilities to the llama.cpp project through GitHub Security Advisories and MITRE. All advisories were closed by the...

9.2CVSS6.5AI score0.00698EPSS
SaveExploits1References9
Packet Storm News
Packet Storm News
•added 2026/08/10 12:00 a.m.•14 views

Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference

The key-value KV cache is the primary throughput optimization in modern large language model LLM inference, enabling prefix reuse across requests. In multi-tenant deployments this cache is shared across tenants, creating a timing side channel: an adversarial tenant can reconstruct another tenant'...

5.2AI score
SaveExploits0
SUSE CVE
SUSE CVE
•added 2026/08/07 5:05 p.m.•17 views

SUSE CVE-2026-43629

llama.cpp builds b4882 through b9058 contain a heap buffer overflow vulnerability in the KV cache state restore path where the statereaddata function computes write size without overflow checking, allowing attackers with write access to the slotsavepath directory to corrupt heap memory. Attackers...

9.2CVSS6.7AI score0.00698EPSS
SaveExploits0References3
EUVD
EUVD
•added 2026/08/07 12:31 a.m.•20 views

EUVD-2026-54284

llama.cpp builds b4882 through b9058 contain a heap buffer overflow vulnerability in the KV cache state restore path where the statereaddata function computes write size without overflow checking, allowing attackers with write access to the slotsavepath directory to corrupt heap memory. Attackers...

9.2CVSS6.7AI score0.00698EPSS
SaveExploits0References2
Rows per page
Query Builder