Lucene search
+L

1 matches found

The Hacker News
The Hacker News
added 2024/10/23 9:54 a.m.24 views

Researchers Reveal 'Deceptive Delight' Method to Jailbreak AI Models

Cybersecurity researchers have shed light on a new adversarial technique that could be used to jailbreak large language models LLMs during the course of an interactive conversation by sneaking in an undesirable instruction between benign ones. The approach has been codenamed Deceptive Delight by...

7.1AI score
SaveExploits0
Rows per page
Query Builder