Lucene search
+L

363 matches found

Kitploit
Kitploit
•added 2026/10/07 4:10 a.m.•10 views

Defenses-for-Tool-Integrated-LLM

面向工具集成式 LLM 智能体的通用防御:抵御对抗性攻击 本仓库包含我们项目的代码与实验,该项目旨在防御工具集成式大型语言模型(LLM)智能体免受对抗性攻击。 概述 我们基于 Agent Security Bench(ASB)构建,评估集成工具与结构化推理(例如思维链、反思)如何影响 LLM 智能体在多种任务场景下对对抗性提示的脆弱性。 本仓库包含: 新的防御策略(例如基于工具的过滤、CoT+Reflection) 改编自 ASB 的攻击场景 实验脚本 基于 Agent Security Bench(ASB) 本项目改编并扩展 了官方 ASB 仓库的代码: Agent Security...

6.2AI score
SaveExploits0References1
Packet Storm News
Packet Storm News
•added 2026/09/28 12:00 a.m.•12 views

Sustained Participation As a Security Resource: The Bounded Participation Channel

Can sustained, per-identity participation be engineered into a security resource? Most anti-Sybil defenses price identity creation rather than identity survival. Once admitted, an adversary may sustain many identities without paying a recurring cost. We introduce the Bounded Participation Channel...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/28 12:00 a.m.•11 views

Distillation Defenses Easily Break after Reinforcement Learning

Distillation attacks copy the reasoning capabilities of closed-source large language models, allowing bad actors to replicate state-of-the-art performance at low cost. Attackers systematically collect a large volume of frontier model reasoning traces and then train i.e., "distill" their own model...

5.8AI score
SaveExploits0
Positive Technologies
Positive Technologies
•added 2026/09/10 12:00 a.m.•24 views

PT-2026-89452

Name of the Vulnerable Software and Affected Versions Traefik versions 3.2.0 through 3.7.12 Description Traefik entrypoint defenses aliasHeadersStrategy, underscoreHeadersStrategy, and forwardedHeaders inspect req.Header but fail to inspect req.Trailer. This allows an unauthenticated client to...

10CVSS7.3AI score0.01573EPSS
SaveExploits4References110
Wired Threat Level
Wired Threat Level
•added 2026/09/01 8:00 p.m.•15 views

OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities

The company will give select partners early access to its Astra AI model—so they have time to shore up their defenses...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/28 12:00 a.m.•8 views

LongPIBench: A Long-Context Benchmark for Prompt Injection

Prompt injection attacks pose a serious security risk to large language models in real-world applications. However, existing prompt injection benchmarks primarily focus on short-context inputs, leaving the attacks and defenses in long-context settings largely unexplored. This gap leads to a...

5.9AI score
SaveExploits0
Akamai Blog
Akamai Blog
•added 2026/08/26 1:00 p.m.•19 views

Not the Coyote, but the Road Runner: The Reality of Autonomous AI Attacks

Autonomous AI security threats aren't novel super-weapons. They're relentless, low-tech attacks that never stop. Learn why traditional defenses fail...

5.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/05 12:00 a.m.•19 views

Hardware Design and Security in the Era of Chiplets and LLMs

The semiconductor industry is undergoing a dual revolution: the shift toward heterogeneous 2.5D chiplet systems and the integration of Large Language Models LLMs into Electronic Design Automation EDA flows. While these paradigms offer unprecedented benefits in yield, modularity, design...

5.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/04 12:00 a.m.•36 views

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personalization, this process concentrates fragmented personal signals, amplifies their impact through reuse, and challenges defenses designed for individual...

5.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/02 12:00 a.m.•33 views

Decoy Images Amplify Caption-Mediated Defenses against Encoded Jailbreaks

We report a counter-intuitive interaction between image inputs and existing black-box defenses on Vision--Language Models VLMs: pairing an encoded jailbreak prompt with an unrelated decoy image can sharply lower attack success rate ASR. The operative change is in the defense pipeline, not in the...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/30 12:00 a.m.•14 views

Blockchain Transaction Simulation Phishing

Cryptocurrency users have increasingly become targets of phishing and scam attacks. To mitigate these threats, leading crypto wallets e.g., MetaMask have introduced transaction simulation, which previews a transaction's balance changes before on-chain execution. While effective against traditiona...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/30 12:00 a.m.•17 views

Temporal Poisoning: Clean-Label Backdoors Via Event Redistribution in SNNs

Backdoor attacks on Spiking Neural Networks SNNs have primarily assumed dirty-label poisoning, in which triggered training samples are relabeled to an attacker-selected class. We study clean-label temporal poisoning, where a fixed timestamp transformation is applied only to the target-class...

5.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/27 12:00 a.m.•18 views

When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-Offs

Jailbreak defenses are essential for protecting large language models LLMs, but they can also introduce secondary costs that weaken model utility. We present a systematic study of these defense trade-offs along three dimensions: performance impact, over-refusal on benign inputs, and inference cos...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/27 12:00 a.m.•98 views

ALIBI: Adaptive Agentic Attacks on LLM-Based Vulnerability Detectors Via Adversarial Code Comments

Large language models are increasingly deployed for security-sensitive tasks such as vulnerability detection and code review. Their reliance on natural-language context embedded in source code exposes a previously underexplored attack surface: adversarial comments that can influence a detector's...

5.9AI score
SaveExploits0
GithubExploit
GithubExploit
•added 2026/07/20 10:50 p.m.•93 views

llm-redteam-helixpay

Red-teaming an LLM support agent — HelixPay A self-contained...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/20 12:00 a.m.•27 views

Broken Gates: Re-Evaluating Web Bot Defenses in the Age of LLM Agents

LLM-based browser agents are rapidly changing the threat landscape for web security. Unlike traditional automation frameworks that execute predefined scripts, these agents can autonomously navigate websites, reason about page content, and interact with web interfaces using natural-language...

5.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/19 12:00 a.m.•62 views

Shared Vulnerabilities in Robustness-Optimized Defenses: One Breach Exposes the Family

Adversarial robustness optimization aims to preserve correct prediction under adversarial perturbations, and has produced substantial robustness gains through methods such as adversarial training and adversarial purification. However, we identify a new security risk: these gains can create shared...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/16 12:00 a.m.•14 views

MemPoison: Uncovering Persistent Memory Threats and Structural Blind Spots in LLM Agents

Persistent external memory enhances agent continuity but introduces persistent security vulnerabilities: adversarial content can be injected via standard interaction channels, retained across turns, and later distort downstream behavior. To address this challenge, we propose MemPoison, a...

5.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/11 12:00 a.m.•14 views

Large Language Models in Misinformation Ecosystems: Misuse, Defense, and Vulnerability

Large language models LLMs have transformed misinformation from a primarily content-centric problem into a broader ecosystem-level security challenge. When misused, LLMs create risks beyond false content generation, enabling attacks on the social contexts, evidence sources, retrieval corpora, and...

6.1AI score
SaveExploits0
The Hacker News
The Hacker News
•added 2026/07/09 10:43 a.m.•39 views

GodDamn Ransomware Uses PoisonX Driver to Disable Endpoint Defenses

Cybersecurity researchers have flagged a new ransomware family called GodDamn that employs the PoisonX kernel driver to neutralize security software as part of its defense evasion strategy. According to a new report published by the Threat Hunter Team from Symantec, the ransomware was first...

6AI score
SaveExploits0
Rows per page
Query Builder