Lucene search
+L

35 matches found

Kitploit
Kitploit
•added 2026/10/05 11:02 a.m.•12 views

KidnapRAG

KidnapRAG: A Black-Box Attack for Hijacking Reasoning in Agentic Retrieval-Augmented Generation Systems 🎓 Paper | 📄 Datasets | 🚀 Quick Start Quick Start Run the following commands from the parent directory of the cloned KidnapRAG repository. cd KidnapRAG conda create -n KidnapRAG python=3.10 cond...

6.3AI score
SaveExploits0References1
Kitploit
Kitploit
•added 2026/10/05 7:42 a.m.•12 views

Adaptive_Greedy_Local_Search

Adaptive Greedy Local Search AGLS Secuestro de Prompts con Preservación Semántica: Un Ataque Adversarial de Caja Negra sobre la Optimización Automática de Prompts ICME 2026 Resumen: Los Modelos de Lenguaje de Gran Tamaño LLMs están cada vez más equipados con módulos automáticos de optimización de...

6.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/27 12:00 a.m.•12 views

DeMark: A Query-Free Black-Box Attack for Quality-Preserving Audio Watermark Removal

Audio watermarking protects digital speech by embedding imperceptible signals for ownership verification and misuse tracing. However, the security of learning-based watermarking remains insufficiently understood under realistic adversarial removal, where attackers cannot access or query the...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/25 12:00 a.m.•11 views

JevAdvBench: A Benchmark and Black-Box Attacks for Reinforcement Learning for Calibrated Decisions Models

Models trained with reinforcement learning for calibrated decisions RLCD, such as Jev, answer a typed question about an input, the state, with a probability, a choice, or a score, and software acts on the answer without a person reading it. Their robustness has not been measured: adversarial...

5.7AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/18 12:00 a.m.•9 views

CASCADE against Jailbreaks: Combination across Stages with Controlled Attack-Defense Evaluation

Defenses against jailbreak attacks on Large Language Models LLMs operate at different pipeline stages, such as input modification or output guard, but it remains unclear which defenses to deploy at each stage and how to combine them. Prior empirical studies, fragmented by inconsistent...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/09 12:00 a.m.•8 views

What Makes Adversarial Examples Transfer across Deepfake Detectors?

Deepfake detectors remain vulnerable to transfer-based black-box attacks, in which adversarial examples are generated on a source surrogate model and transferred to a target model, unknown to the attacker. Yet how source--target compatibility shapes attack success remains poorly understood. Prior...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/01 12:00 a.m.•19 views

Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models

Text-to-Speech TTS foundation models are increasingly fine-tuned on private datasets to synthesize highly personalized voices, introducing severe privacy risks by exposing both biometric identities and sensitive speech content. Existing black-box membership inference attacks MIAs follow a two-sta...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/28 12:00 a.m.•8 views

REPLICANT: Learning Policies for Evading and Hardening Malware Detectors

To determine the real-world effectiveness of machine learning based malware detection, it is vital to evaluate its robustness against highly capable adversaries. However, state-of-the-art attacks do not effectively model realistic adversaries, as they often assume access to privileged information...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/27 12:00 a.m.•34 views

Daydreaming: Stealing Hidden Agent Skills through Black-Box Task Interaction

Agent skills bundle instructions, reference data, and executable helpers that let a general agent perform specialized tasks. Hosted providers can keep these files secret while selling access to task results, making the skill itself a valuable target. Existing disclosure defenses can block request...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/27 12:00 a.m.•41 views

RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution

LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbreaks can trigger harmful tool use and persistent state changes, creating greater risks than unsafe text generation alone. Existing automatic red-teaming methods often rely on fixed attacks, while recent...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/04 12:00 a.m.•47 views

Behavioral Skill Reconstruction: Reconstructing Hidden Functionality from LLM Agent Skills

Closed source agent skills may encode proprietary instructions, scripts, constants, and data. Providers may offer their capabilities as services while keeping the underlying packages hidden. Prior work focuses on prompt injection attacks that directly disclose these artifacts, and existing defens...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/27 12:00 a.m.•97 views

ALIBI: Adaptive Agentic Attacks on LLM-Based Vulnerability Detectors Via Adversarial Code Comments

Large language models are increasingly deployed for security-sensitive tasks such as vulnerability detection and code review. Their reliance on natural-language context embedded in source code exposes a previously underexplored attack surface: adversarial comments that can influence a detector's...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/13 12:00 a.m.•75 views

HarmQ: Harmonic Backdoor Attacks against Quantum Neural Networks

Quantum Neural Networks QNNs have emerged as a promising paradigm for quantum machine learning in the Noisy Intermediate-Scale Quantum NISQ era, leveraging quantum phenomena such as superposition and entanglement to process information in exponentially large Hilbert spaces. However, QNNs inherit...

6.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/29 12:00 a.m.•19 views

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection

Researchers and practitioners increasingly apply Large Language Models LLMs for automated vulnerability detection. Recent work has shown that LLMs are susceptible to the same cognitive heuristics that bias human judgment. Yet, no work has investigated whether these heuristics affect a model's...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/09 12:00 a.m.•37 views

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However, this paradigm introduces a critical vulnerability: malicious content injected into memory can be persistently recalled and repeatedly influence age...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/01 12:00 a.m.•19 views

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

Diffusion large language models dLLMs generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safety surface distinct from autoregressive LLMs. Because mask tokens are native inputs and tokens are committed by confidence rather than position,...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/05/18 12:00 a.m.•49 views

Babel: Jailbreaking Safety Attention Via Obfuscation Distribution Optimized Sampling

Despite rigorous safety alignment, Large Language Models LLMs remain vulnerable to jailbreak attacks. Existing black-box methods often rely on heuristic templates or exhaustive trials, lacking mechanistic interpretability and query efficiency. In this study, we investigate an intrinsic...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/03/24 12:00 a.m.•20 views

Targeted Adversarial Traffic Generation : Black-Box Approach to Evade Intrusion Detection Systems in IoT Networks

The integration of machine learning ML algorithms into Internet of Things IoT applications has introduced significant advantages alongside vulnerabilities to adversarial attacks, especially within IoT-based intrusion detection systems IDS. While theoretical adversarial attacks have been extensive...

5.7AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/11/24 12:00 a.m.•15 views

Frequency Bias Matters: Diving into Robust and Generalized Deep Image Forgery Detection

As deep image forgery powered by AI generative models, such as GANs, continues to challenge today's digital world, detecting AI-generated forgeries has become a vital security topic. Generalizability and robustness are two critical concerns of a forgery detector, determining its reliability when...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/11/20 12:00 a.m.•17 views

"To Survive, I Must Defect": Jailbreaking LLMs Via the Game-Theory Scenarios

As LLMs become more common, non-expert users can pose risks, prompting extensive research into jailbreak attacks. However, most existing black-box jailbreak attacks rely on hand-crafted heuristics or narrow search spaces, which limit scalability. Compared with prior attacks, we propose Game-Theor...

6.9AI score
SaveExploits0
Rows per page
Query Builder