Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2026/04/21 12:0 a.m.6 views

Involuntary In-Context Learning: Exploiting Few-Shot Pattern Completion to Bypass Safety Alignment in GPT-5.4

Safety alignment in large language models relies on behavioral training that can be overridden when sufficiently strong in-context patterns compete with learned refusal behaviors. We introduce Involuntary In-Context Learning IICL, an attack class that uses abstract operator framing with few-shot...

5.7AI score
SaveExploits0
Rows per page
Query Builder