Lucene search
+L

110 matches found

Packet Storm News
Packet Storm News
added 2025/09/19 12:00 a.m.12 views

Automated Cyber Defense with Generalizable Graph-Based Reinforcement Learning Agents

Deep reinforcement learning RL is emerging as a viable strategy for automated cyber defense ACD. The traditional RL approach represents networks as a list of computers in various states of safety or threat. Unfortunately, these models are forced to overfit to specific network topologies, renderin...

7.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/26 12:00 a.m.9 views

DRMD: Deep Reinforcement Learning for Malware Detection under Concept Drift

Malware detection in real-world settings must deal with evolving threats, limited labeling budgets, and uncertain predictions. Traditional classifiers, without additional mechanisms, struggle to maintain performance under concept drift in malware domains, as their supervised learning formulation...

6.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/26 12:00 a.m.12 views

Attackers Strike Back? Not Anymore -- an Ensemble of RL Defenders Awakens for APT Detection

Advanced Persistent Threats APTs represent a growing menace to modern digital infrastructure. Unlike traditional cyberattacks, APTs are stealthy, adaptive, and long-lasting, often bypassing signature-based detection systems. This paper introduces a novel framework for APT detection that unites de...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/20 12:00 a.m.8 views

Aura-CAPTCHA: a Reinforcement Learning and GAN-Enhanced Multi-Modal CAPTCHA System

Aura-CAPTCHA was developed as a multi-modal CAPTCHA system to address vulnerabilities in traditional methods that are increasingly bypassed by AI technologies, such as Optical Character Recognition OCR and adversarial image processing. The design integrated Generative Adversarial Networks GANs fo...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/14 12:00 a.m.9 views

REFN: a Reinforcement-Learning-From-Network Framework against 1-Day/N-Day Exploitations

The exploitation of 1 day or n day vulnerabilities poses severe threats to networked devices due to massive deployment scales and delayed patching average Mean Time To Patch exceeds 60 days. Existing defenses, including host based patching and network based filtering, are inadequate due to limite...

7.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/12 12:00 a.m.8 views

Attacks and Defenses against LLM Fingerprinting

As large language models are increasingly deployed in sensitive environments, fingerprinting attacks pose significant privacy and security risks. We present a study of LLM fingerprinting from both offensive and defensive perspectives. Our attack methodology uses reinforcement learning to...

6.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/07 12:00 a.m.8 views

RL-MoE: an Image-Based Privacy Preserving Approach in Intelligent Transportation System

The proliferation of AI-powered cameras in Intelligent Transportation Systems ITS creates a severe conflict between the need for rich visual data and the fundamental right to privacy. Existing privacy-preserving mechanisms, such as blurring or encryption, are often insufficient, creating an...

6.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/05 12:00 a.m.19 views

When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs

As large language models become increasingly integrated into daily life, audio has emerged as a key interface for human-AI interaction. However, this convenience also introduces new vulnerabilities, making audio a potential attack surface for adversaries. Our research introduces WhisperInject, a...

7.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/04 12:00 a.m.14 views

Secure MmWave Beamforming with Proactive-ISAC Defense against Beam-Stealing Attacks

Millimeter-wave mmWave communication systems face increasing susceptibility to advanced beam-stealing attacks, posing a significant physical layer security threat. This paper introduces a novel framework employing an advanced Deep Reinforcement Learning DRL agent for proactive and adaptive defens...

6.7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/29 12:00 a.m.7 views

Secure Tug-Of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security

The rapid advancement of multimodal large language models MLLMs has led to breakthroughs in various applications, yet their security remains a critical challenge. One pressing issue involves unsafe image-query pairs--jailbreak inputs specifically designed to bypass security constraints and elicit...

7.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/25 12:00 a.m.11 views

PurpCode: Reasoning for Safer Code Generation

We introduce PurpCode, the first post-training recipe for training safe code reasoning models towards generating secure code and defending against malicious cyberactivities. PurpCode trains a reasoning model in two stages: i Rule Learning, which explicitly teaches the model to reference cybersafe...

7.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/17 12:00 a.m.10 views

Thought Purity: Defense Paradigm for Chain-Of-Thought Attack

While reinforcement learning-trained Large Reasoning Models LRMs, e.g., Deepseek-R1 demonstrate advanced reasoning capabilities in the evolving Large Language Models LLMs domain, their susceptibility to security threats remains a critical vulnerability. This weakness is particularly evident in...

7.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/10 12:00 a.m.24 views

Agent Safety Alignment Via Reinforcement Learning

The emergence of autonomous Large Language Model LLM agents capable of tool usage has introduced new safety risks that go beyond traditional conversational misuse. These agents, empowered to execute external functions, are vulnerable to both user-initiated threats e.g., adversarial prompts and...

7.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/07 12:00 a.m.11 views

Beyond Training-Time Poisoning: Component-Level and Post-Training Backdoors in Deep Reinforcement Learning

Deep Reinforcement Learning DRL systems are increasingly used in safety-critical applications, yet their security remains severely underexplored. This work investigates backdoor attacks, which implant hidden triggers that cause malicious actions only when specific inputs appear in the observation...

7.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/06 12:00 a.m.11 views

Adaptive Malware Detection Using Sequential Feature Selection: a Dueling Double Deep Q-Network (D3QN) Framework for Intelligent Classification

Traditional malware detection methods exhibit computational inefficiency due to exhaustive feature extraction requirements, creating accuracy-efficiency trade-offs that limit real-time deployment. We formulate malware classification as a Markov Decision Process with episodic feature acquisition a...

6.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/27 12:00 a.m.10 views

ARMOR: Robust Reinforcement Learning-Based Control for UAVs under Physical Attacks

Unmanned Aerial Vehicles UAVs depend on onboard sensors for perception, navigation, and control. However, these sensors are susceptible to physical attacks, such as GPS spoofing, that can corrupt state estimates and lead to unsafe behavior. While reinforcement learning RL offers adaptive control...

6.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/24 12:00 a.m.12 views

Autonomous Cyber Resilience Via a Co-Evolutionary Arms Race within a Fortified Digital Twin Sandbox

The convergence of IT and OT has created hyper-connected ICS, exposing critical infrastructure to a new class of adaptive, intelligent adversaries that render static defenses obsolete. Existing security paradigms often fail to address a foundational "Trinity of Trust," comprising the fidelity of...

7.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/23 12:00 a.m.11 views

Adaptive Alert Prioritisation in Security Operations Centres Via Learning to Defer with Human Feedback

Alert prioritisation AP is crucial for security operations centres SOCs to manage the overwhelming volume of alerts and ensure timely detection and response to genuine threats, while minimising alert fatigue. Although predictive AI can process large alert volumes and identify known patterns, it...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/22 12:00 a.m.12 views

VulStamp: Vulnerability Assessment Using Large Language Model

Although modern vulnerability detection tools enable developers to efficiently identify numerous security flaws, indiscriminate remediation efforts often lead to superfluous development expenses. This is particularly true given that a substantial portion of detected vulnerabilities either possess...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/22 12:00 a.m.13 views

LLM-Based Dynamic Differential Testing for Database Connectors with Reinforcement Learning-Guided Prompt Selection

Database connectors are critical components enabling applications to interact with underlying database management systems DBMS, yet their security vulnerabilities often remain overlooked. Unlike traditional software defects, connector vulnerabilities exhibit subtle behavioral patterns and are...

7.2AI score
SaveExploits0
Rows per page
Query Builder