Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/05/16 12:0 a.m.8 views

LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs

Efficient red-teaming method to uncover vulnerabilities in Large Language Models LLMs is crucial. While recent attacks often use LLMs as optimizers, the discrete language space make gradient-based methods struggle. We introduce LARGO Latent Adversarial Reflection through Gradient Optimization, a...

7AI score
SaveExploits0
Rows per page
Query Builder