Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2025/10/06 12:0 a.m.8 views

Imperceptible Jailbreaking against Large Language Models

Jailbreaking attacks on the vision modality typically rely on imperceptible adversarial perturbations, whereas attacks on the textual modality are generally assumed to require visible modifications e.g., non-semantic suffixes. In this paper, we introduce imperceptible jailbreaks that exploit a...

7.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/09/17 12:0 a.m.4 views

LLM Jailbreak Detection for (Almost) Free!

Large language models LLMs enhance security through alignment when widely used, but remain susceptible to jailbreak attacks capable of producing inappropriate content. Jailbreak detection methods show promise in mitigating jailbreak attacks through the assistance of other models or multiple model...

6.8AI score
SaveExploits0
Rows per page
Query Builder