Lucene search
+L

117 matches found

Kitploit
Kitploit
•added 2026/10/08 6:21 a.m.•12 views

Pesidious

Malware Mutation using Deep Reinforcement Learning and GANs The purpose of the tool is to use artificial intelligence to mutate a malware PE32 only sample to bypass AI powered classifiers while keeping its functionality intact. In the past, notable work has been done in this domain with researche...

5.8AI score
SaveExploits0References8
Kitploit
Kitploit
•added 2026/10/07 11:45 p.m.•22 views

AutoPentest-DRL

AutoPentest-DRL: Automated Penetration Testing Using Deep Reinforcement Learning AutoPentest-DRL is an automated penetration testing framework based on Deep Reinforcement Learning DRL techniques. AutoPentest-DRL can determine the most appropriate attack path for a given logical network, and can...

6.5AI score
SaveExploits0References4
Kitploit
Kitploit
•added 2026/10/07 11:16 p.m.•11 views

GuardReasoner-VL

GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning Yue Liu, Shengfang Zhai, Mingzhe Du Yulin Chen, Tri Cao, Hongcheng Gao, Cheng Wang Xinfeng Li, Kun Wang, Junfeng Fang, Jiaheng Zhang, Bryan Hooi 1National University of Singapore, 2Nanyang Technological University To enhance the safety ...

6.2AI score
SaveExploits0References11
Kitploit
Kitploit
•added 2026/10/07 10:26 p.m.•21 views

CyberBattleSim

CyberBattleSim April 8th, 2021: See the announcement on the Microsoft Security Blog. CyberBattleSim is an experimentation research platform to investigate the interaction of automated agents operating in a simulated abstract enterprise network environment. The simulation provides a high-level...

6.3AI score
SaveExploits0References3
Kitploit
Kitploit
•added 2026/10/07 10:06 p.m.•13 views

PISmith

PIForge An Open Framework for RL-based Prompt Injection Red Teaming 💻 Code · 🤗 Models · 📜 Papers PIForge is the shared codebase for PISmith and Climbing the Hill. PISmith addresses the sparse reward problem in prompt injection red teaming, helping RL attackers explore and learn from rare successf...

6.3AI score
SaveExploits0References10
Kitploit
Kitploit
•added 2026/10/07 9:01 p.m.•16 views

StealthRL

StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors Paper arXiv Demo Model Hugging Face Benchmark Dataset Hugging Face Abstract AI-text detectors are increasingly used in high-stakes settings, yet their robustness to meaning-preserving adversarial...

6.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/05 12:00 a.m.•5 views

Understanding and Enhancing Backdoor Persistency in LLM Agent Post-Training

Developers can build LLM agents by adapting third-party models through benign post-training. We study a supply-chain threat in which an attacker supplies a model with a backdoor: hidden behavior that produces malicious outputs when a particular input pattern appears. Focusing on...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/01 12:00 a.m.•11 views

Homomorphic Advantage Operator: Stabilizing Reinforcement Learning under Fully Homomorphic Encryption Constraints

Privacy-preserving machine learning presents significant deployment challenges on the cloud for intelligent systems with confidential data. Fully Homomorphic Encryption FHE offers a compelling solution for secure computation, preserving data confidentiality of cloud computations. However, applyin...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/01 12:00 a.m.•10 views

Learnt Attacks on Quantum Key Distribution under Channel Noise and Device Drift

Quantum key distribution QKD links are provisioned from security analyses of stationary channels, whereas the devices that determine the channel drift between recalibrations. Whether an eavesdropper who cannot alter the channel's own noise gains by following that drift has not been quantified...

5.9AI score
SaveExploits0
NVD
NVD
•added 2026/09/30 3:22 p.m.•14 views

CVE-2026-103270

LightLLM through 1.2.0 mounts reinforcement learning control routes on the public HTTP API without authentication checks. Unauthenticated attackers can call endpoints like /pausegeneration, /abortrequest, /flushcache, and /initweightsupdategroup to disrupt inference operations and wedge workers o...

8.7CVSS0.00513EPSS
SaveExploits0References5
Rapid7 Vulnerability Database (full)
Rapid7 Vulnerability Database (full)
•added 2026/09/30 12:00 a.m.•15 views

CVE-2026-103270: Missing Authentication for Critical Function

LightLLM through 1.2.0 mounts reinforcement learning control routes on the public HTTP API without authentication checks. Unauthenticated attackers can call endpoints like /pausegeneration, /abortrequest, /flushcache, and /initweightsupdategroup to disrupt inference operations and wedge workers o...

8.7CVSS5.9AI score0.00513EPSS
SaveExploits0References2
Packet Storm News
Packet Storm News
•added 2026/09/30 12:00 a.m.•12 views

Towards Hierarchical Cyber Defense with Large Language Models: From Planning to Execution

An autonomous cyber defender trained with reinforcement learning RL is typically tied to the network on which it was trained, limiting its ability to generalize as network scale changes. Hierarchical RL reduces decision complexity by separating strategic targeting from tactical execution, but it...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/27 12:00 a.m.•10 views

RAISE: Reinforcing Access Control Policy Synthesis in LLMs Via Symbolic Evaluation

Translating natural-language access-control requirements into policies requires careful reasoning about permissions, constraints, and exceptions, and even frontier LLMs often produce policies that violate the intended authorization semantics. We construct CedarInstruct, to our knowledge the first...

5.8AI score
SaveExploits0
The Hacker News
The Hacker News
•added 2026/08/19 6:06 p.m.•26 views

OpenAI Pauses Frontier RL Training as It Tightens Defenses Against Unsafe AI Behavior

OpenAI on Tuesday revealed that it paused reinforcement learning RL training for its latest artificial intelligence AI models for two weeks while it shored up additional defenses and increased the scope of its monitoring to avert another Hugging Face-like incident. "As models become more capable,...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/06 12:00 a.m.•34 views

ARTA: Adaptive Reinforcement-Learning-Based Throttling Agent for RowHammer Vulnerabilities

RowHammer vulnerability continues to intensify with DRAM scaling, reducing the activation threshold needed to induce bitflips and rendering existing defenses such as TRR, ECC, and refresh-based mechanisms vulnerable to sophisticated multi-bank hammering patterns. This work presents ARTA, a...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/05/16 12:00 a.m.•22 views

A Red Teaming Framework for Evaluating Robustness of AI-Enabled Security Orchestration, Automation, and Response Systems

AI-enabled Security Orchestration, Automation, and Response SOAR systems increasingly employ autonomous agents for cyber defense, yet their resilience to adaptive adversaries is underexplored. We introduce an autonomous red teaming framework that integrates large language models LLMs with...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/05/10 12:00 a.m.•32 views

Operationalizing Cybersecurity Governance for Mitigation Planning with Attack-Path Modeling and Reinforcement Learning

We address a fundamental challenge in cybersecurity operations of translating governance frameworks into actionable mitigation decisions under realistic resource constraints. Frameworks such as the NIST Cybersecurity Framework CSF provide widely adopted measures of organizational maturity, but do...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/05/01 12:00 a.m.•18 views

STARE: Step-Wise Temporal Alignment and Red-Teaming Engine for Multi-Modal Toxicity Attack

Red-teaming Vision-Language Models is essential for identifying vulnerabilities where adversarial image-text inputs trigger toxic outputs. Existing approaches treat image generation as a black box, returning only terminal toxicity scores and leaving open the question of when and how toxic semanti...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/04/30 12:00 a.m.•89 views

XekRung Technical Report

We present XekRung, a frontier large language model for cybersecurity, designed to provide comprehensive security capabilities. To achieve this, we develop diverse data synthesis pipelines tailored to the cybersecurity domain, enabling the scalable construction of high-quality training data and...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/04/23 12:00 a.m.•16 views

Risk Models As Mediating Artifacts: A Postphenomenological Analysis of the CIIM Framework in Cybersecurity Practice

This article applies postphenomenological theory to the field of cybersecurity risk management, arguing that formal risk models function as mediating artifacts that shape how security practitioners or analysts perceive, interpret, and act on threats. Based on Don Ihde's taxonomy on human-technolo...

5.3AI score
SaveExploits0
Rows per page
Query Builder