Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/05/14 12:0 a.m.7 views

Adversarial Attack on Large Language Models Using Exponentiated Gradient Descent

As Large Language Models LLMs are widely used, understanding them systematically is key to improving their safety and realizing their full potential. Although many models are aligned using techniques such as reinforcement learning from human feedback RLHF, they are still vulnerable to jailbreakin...

7.4AI score
SaveExploits0
Rows per page
Query Builder