Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/11/24 12:0 a.m.4 views

Defending Large Language Models against Jailbreak Exploits with Responsible AI Considerations

Large Language Models LLMs remain susceptible to jailbreak exploits that bypass safety filters and induce harmful or unethical behavior. This work presents a systematic taxonomy of existing jailbreak defenses across prompt-level, model-level, and training-time interventions, followed by three...

7.3AI score
SaveExploits0
Rows per page
Query Builder