Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2026/04/13 12:0 a.m.8 views

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs

Large Language Models LLMs are increasingly deployed across diverse domains, yet their vulnerability to jailbreak attacks, where adversarial inputs bypass safety mechanisms to elicit harmful outputs, poses significant security risks. While prior work has primarily focused on prompt injection...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2026/02/04 12:0 a.m.9 views

Inference-Time Backdoors Via Hidden Instructions in LLM Chat Templates

Open-weight language models are increasingly used in production settings, raising new security challenges. One prominent threat in this context is backdoor attacks, in which adversaries embed hidden behaviors in language models that activate under specific conditions. Previous work has assumed th...

5.5AI score
SaveExploits0
Rows per page
Query Builder