Lucene search
+L

107 matches found

Kitploit
Kitploit
тАвadded 2026/09/06 6:48 p.m.тАв3 views

dataset

ЁЯЪА CySecBench: Generative AI рдкрд░ рдЖрдзрд╛рд░рд┐рдд рд╕рд╛рдЗрдмрд░рд╕реБрд░рдХреНрд╖рд╛-рдХреЗрдВрджреНрд░рд┐рдд рдкреНрд░реЙрдореНрдкреНрдЯ рдбреЗрдЯрд╛рд╕реЗрдЯ рдмрдбрд╝реЗ рднрд╛рд╖рд╛ рдореЙрдбрд▓реЛрдВ рдХреЗ рдорд╛рдирдХреАрдХрд░рдг рдХреЗ рд▓рд┐рдП ЁЯЫбя╕П рдмрдбрд╝реЗ рднрд╛рд╖рд╛ рдореЙрдбрд▓реЛрдВ рдХреЗ рдорд╛рдирдХреАрдХрд░рдг рдХреЗ рд▓рд┐рдП рд╕рдмрд╕реЗ рдмрдбрд╝рд╛ рдФрд░ рд╕рдмрд╕реЗ рд╡реНрдпрд╛рдкрдХ Generative AI-рдЖрдзрд╛рд░рд┐рдд рд╕рд╛рдЗрдмрд░рд╕реБрд░рдХреНрд╖рд╛-рдХреЗрдВрджреНрд░рд┐рдд рдбреЗрдЯрд╛рд╕реЗрдЯ ЁЯМЯ рдЕрд╡рд▓реЛрдХрди CySecBench рд╢реЛрдзрдкрддреНрд░ рдкреНрд░рд╕реНрддреБрдд рдХрд░рддрд╛ рд╣реИ: ЁЯОп рдПрдХ рдЕрддреНрдпрд╛рдзреБрдирд┐рдХ рдбреЗрдЯрд╛рд╕реЗрдЯ...

6AI score
SaveExploits0References12
Kitploit
Kitploit
тАвadded 2026/09/06 11:07 a.m.тАв7 views

ai-llm-red-team-handbook

AI / LLM Red Team рдлреАрд▓реНрдб рдореИрдиреБрдЕрд▓ рдФрд░ рд╕рд▓рд╛рд╣рдХрд╛рд░ рд╣реИрдВрдбрдмреБрдХ рдмрдбрд╝реЗ рднрд╛рд╖рд╛ рдореЙрдбрд▓ LLM, AI рдПрдЬреЗрдВрдЯ, RAG рдкрд╛рдЗрдкрд▓рд╛рдЗрди рдФрд░ AI-рд╕рдХреНрд╖рдо рдЕрдиреБрдкреНрд░рдпреЛрдЧреЛрдВ рдкрд░ AI/LLM рд░реЗрдб рдЯреАрдо рдореВрд▓реНрдпрд╛рдВрдХрди рдХрд░рдиреЗ рдХреЗ рд▓рд┐рдП рдПрдХ рд╡реНрдпрд╛рдкрдХ рд╕рдВрдЪрд╛рд▓рдирд╛рддреНрдордХ рдЯреВрд▓рдХрд┐рдЯред рдпрд╣ рд░рд┐рдкреЙрдЬрд┐рдЯрд░реА рд╕рд╛рдорд░рд┐рдХ рдХреНрд╖реЗрддреНрд░ рдорд╛рд░реНрдЧрджрд░реНрд╢рди рдФрд░ рд░рдгрдиреАрддрд┐рдХ рдкрд░рд╛рдорд░реНрд╢ рдврд╛рдБрдЪреЗ рджреЛрдиреЛрдВ рдкреНрд░рджрд╛рди рдХрд░рддреА рд╣реИред ЁЯУЦ GitBook рдиреЗрд╡рд┐рдЧреЗрд╢рди: рдкреВрд░реН...

5.9AI score
SaveExploits0References1
Kitploit
Kitploit
тАвadded 2026/09/06 4:05 a.m.тАв2 views

PS4-5.05-Kernel-Exploit

PS4 5.05 рдХрд░реНрдирд▓ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рд╕рд╛рд░рд╛рдВрд╢ рдЗрд╕ рдкрд░рд┐рдпреЛрдЬрдирд╛ рдореЗрдВ рдЖрдкрдХреЛ PlayStation 4 рдХреЗ 5.05 рдкрд░ рджреВрд╕рд░реЗ "bpf" рдХрд░реНрдирд▓ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рдХрд╛ рдкреВрд░реНрдг рдХрд╛рд░реНрдпрд╛рдиреНрд╡рдпрди рдорд┐рд▓реЗрдЧрд╛ред рдпрд╣ рдЖрдкрдХреЛ рдХрд░реНрдирд▓ рдХреЗ рд░реВрдк рдореЗрдВ рдордирдорд╛рдирд╛ рдХреЛрдб рдЪрд▓рд╛рдиреЗ рдХреА рдЕрдиреБрдорддрд┐ рджреЗрдЧрд╛, рддрд╛рдХрд┐ рд╕рд┐рд╕реНрдЯрдо рдореЗрдВ рдЬреЗрд▓рдмреНрд░реЗрдХрд┐рдВрдЧ рдФрд░ рдХрд░реНрдирд▓-рд╕реНрддрд░реАрдп рд╕рдВрд╢реЛрдзрди рдХрд┐рдП рдЬрд╛ рд╕рдХреЗрдВред рдЗрд╕ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рдореЗрдВ Mira рдФрд░ Vortex рдХреЗ HE...

5.9AI score
SaveExploits0References4
Kitploit
Kitploit
тАвadded 2026/09/04 5:29 p.m.тАв6 views

PS4-5.05-Kernel-Exploit

PS4 5.05 рдХрд░реНрдиреЗрд▓ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рд╕рд╛рд░рд╛рдВрд╢ рдЗрд╕ рдкреНрд░реЛрдЬреЗрдХреНрдЯ рдореЗрдВ рдЖрдкрдХреЛ PlayStation 4 рдкрд░ 5.05 рдХреЗ рд▓рд┐рдП рджреВрд╕рд░реЗ "bpf" рдХрд░реНрдиреЗрд▓ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рдХрд╛ рдкреВрд░реНрдг рдХрд╛рд░реНрдпрд╛рдиреНрд╡рдпрди рдорд┐рд▓реЗрдЧрд╛ред рдпрд╣ рдЖрдкрдХреЛ рдХрд░реНрдиреЗрд▓ рдХреЗ рд░реВрдк рдореЗрдВ рдордирдорд╛рдирд╛ рдХреЛрдб рдЪрд▓рд╛рдиреЗ рдХреА рдЕрдиреБрдорддрд┐ рджреЗрдЧрд╛, рддрд╛рдХрд┐ рд╕рд┐рд╕реНрдЯрдо рдореЗрдВ рдЬреЗрд▓рдмреНрд░реЗрдХрд┐рдВрдЧ рдФрд░ рдХрд░реНрдиреЗрд▓-рд╕реНрддрд░реАрдп рд╕рдВрд╢реЛрдзрди рдХрд┐рдП рдЬрд╛ рд╕рдХреЗрдВред рдЗрд╕ рдПрдХреНрд╕рдкреНрд▓реЙрдЗрдЯ рдореЗрдВ Mira рдФрд░...

5.9AI score
SaveExploits0References3
Packet Storm News
Packet Storm News
тАвadded 2026/06/05 12:00 a.m.тАв28 views

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (And Fail) Red Team Attacks

Standard AI red teaming evaluations reduce adversarial campaigns to a single binary outcome, attack success rate ASR, not taking into account the sequential structure of how models resist or yield to attacks. We propose applying process mining, a discipline for discovering and analyzing process...

5.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/06/04 12:00 a.m.тАв130 views

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense

Despite advances in safety alignment, large language models remain vulnerable to continuously evolving jailbreaks. Existing fine-tuned safety classifiers cannot adapt to these evolving attacks, while adaptive memory-based guardrails tend to over-refuse benign queries that resemble stored attacks...

5.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/06/01 12:00 a.m.тАв14 views

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

Diffusion large language models dLLMs generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safety surface distinct from autoregressive LLMs. Because mask tokens are native inputs and tokens are committed by confidence rather than position,...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/27 12:00 a.m.тАв16 views

Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking

Jailbreak attacks on large language models LLMs aim to induce LLMs to produce content that they are expected to refuse. Automated black-box jailbreak generation is especially important for safety evaluation, where the attacker observes only model outputs and needs to automatically search for...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/26 12:00 a.m.тАв32 views

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

Large Language Models LLMs are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulting in harmful or inappropriate outputs. Such attacks, including jailbreaking and prompt injection, pose significant risks to the integrity and...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/11 12:00 a.m.тАв21 views

Guaranteed Jailbreaking Defense Via Disrupt-And-Rectify Smoothing

This paper proposes a guaranteed defense method for large language models LLMs to safeguard against jailbreaking attacks. Drawing inspiration from the denoised-smoothing approach in the adversarial defense domain, we propose a novel smoothing-based defense method, termed Disrupt-and-Rectify...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/11 12:00 a.m.тАв14 views

Re-Triggering Safeguards within LLMs for Jailbreak Detection

This paper proposes a jailbreaking prompt detection method for large language models LLMs to defend against jailbreak attacks. Although recent LLMs are equipped with built-in safeguards, it remains possible to craft jailbreaking prompts that bypass them. We argue that such jailbreaking prompts ar...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/11 12:00 a.m.тАв16 views

LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments

The rapid proliferation of LLM-based autonomous agents in real operating system environments introduces a new category of safety risk beyond content safety: behavior jailbreak, where an adversary induces an agent to execute dangerous OS-level operations with irreversible consequences. Existing...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/05/08 12:00 a.m.тАв17 views

Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment

Recent advancements in visual context compression enable MLLMs to process ultra-long contexts efficiently by rendering text into images. However, we identify a critical vulnerability inherent to this paradigm: lowering image resolution inadvertently catalyzes jailbreaking. Our experiments reveal...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/04/27 12:00 a.m.тАв11 views

Jailbreaking Frontier Foundation Models through Intention Deception

Large vision-language models exhibit remarkable capability but remain highly susceptible to jailbreaking. Existing safety training approaches aim to have the model learn a refusal boundary between safe and unsafe, based on the user's intent. It has been found that this binary training regime ofte...

5.3AI score
SaveExploits0
Schneier on Security
Schneier on Security
тАвadded 2026/03/10 9:50 a.m.тАв15 views

Jailbreaking the F-35 Fighter Jet

Countries around the world are becoming increasingly concerned about their dependencies on the US. If you've purchase US-made F-35 fighter jets, you are dependent on the US for software maintenance. The Dutch Defense Secretary recently said that he could jailbreak the planes to accept third-party...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/03/06 12:00 a.m.тАв8 views

Two Frames Matter: A Temporal Attack for Text-To-Video Model Jailbreaking

Recent text-to-video T2V models can synthesize complex videos from lightweight natural language prompts, raising urgent concerns about safety alignment in the event of misuse in the real world. Prior jailbreak attacks typically rewrite unsafe prompts into paraphrases that evade content filters...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/02/27 12:00 a.m.тАв12 views

Jailbreak Foundry: From Papers to Runnable Attacks for Reproducible Benchmarking

Jailbreak techniques for large language models LLMs evolve faster than benchmarks, making robustness estimates stale and difficult to compare across papers due to drift in datasets, harnesses, and judging protocols. We introduce JAILBREAK FOUNDRY JBF, a system that addresses this gap via a...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/02/18 12:00 a.m.тАв16 views

Recursive Language Models for Jailbreak Detection: A Procedural Defense for Tool-Augmented Agents

Jailbreak prompts are a practical and evolving threat to large language models LLMs, particularly in agentic systems that execute tools over untrusted content. Many attacks exploit long-context hiding, semantic camouflage, and lightweight obfuscations that can evade single-pass guardrails. We...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/02/11 12:00 a.m.тАв13 views

Jailbreaking Leaves a Trace: Understanding and Detecting Jailbreak Attacks from Internal Representations of Large Language Models

Jailbreaking large language models LLMs has emerged as a critical security challenge with the widespread deployment of conversational AI systems. Adversarial users exploit these models through carefully crafted prompts to elicit restricted or unsafe outputs, a phenomenon commonly referred to as...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
тАвadded 2026/02/06 12:00 a.m.тАв33 views

TrapSuffix: Proactive Defense against Adversarial Suffixes in Jailbreaking

Suffix-based jailbreak attacks append an adversarial suffix, i.e., a short token sequence, to steer aligned LLMs into unsafe outputs. Since suffixes are free-form text, they admit endlessly many surface forms, making jailbreak mitigation difficult. Most existing defenses depend on passive detecti...

5.3AI score
SaveExploits0
Rows per page
Query Builder