Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
added 2025/11/04 12:0 a.m.7 views

Jailbreaking in the Haystack

Recent advances in long-context language models LMs have enabled million-token inputs, expanding their capabilities across complex tasks like computer-use agents. Yet, the safety implications of these extended contexts remain unclear. To bridge this gap, we introduce NINJA short for...

7.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/05/29 12:0 a.m.7 views

Hijacking Large Language Models Via Adversarial In-Context Learning

In-context learning ICL has emerged as a powerful paradigm leveraging LLMs for specific downstream tasks by utilizing labeled examples as demonstrations demos in the preconditioned prompts. Despite its promising performance, crafted adversarial attacks pose a notable threat to the robustness of...

7.4AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/05/26 12:0 a.m.16 views

One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP

Deep Neural Networks DNNs have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve frequent queries to the target model or rely on surrogate models closely mirroring the target model -- often trained with subsets of the target model's traini...

6.8AI score
SaveExploits0
Rows per page
Query Builder