Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
added 2025/12/05 12:0 a.m.27 views

TeleAI-Safety: A Comprehensive LLM Jailbreaking Benchmark Towards Attacks, Defenses, and Evaluations

While the deployment of large language models LLMs in high-value industries continues to expand, the systematic assessment of their safety against jailbreak and prompt-based attacks remains insufficient. Existing safety evaluation benchmarks and frameworks are often limited by an imbalanced...

7.5AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/18 12:0 a.m.6 views

MAJIC: Markovian Adaptive Jailbreaking Via Iterative Composition of Diverse Innovative Strategies

Large Language Models LLMs have exhibited remarkable capabilities but remain vulnerable to jailbreaking attacks, which can elicit harmful content from the models by manipulating the input prompts. Existing black-box jailbreaking techniques primarily rely on static prompts crafted with a single,...

6.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/04 12:0 a.m.7 views

Large Reasoning Models Are Autonomous Jailbreak Agents

Jailbreaking -- bypassing built-in safety mechanisms in AI models -- has traditionally required complex technical procedures or specialized human expertise. In this study, we show that the persuasive capabilities of large reasoning models LRMs simplify and scale jailbreaking, converting it into a...

7.1AI score
SaveExploits0
Rows per page
Query Builder