Lucene search
+L

1 matches found

Microsoft Secure
Microsoft Secure
added 2026/02/09 5:12 p.m.8 views

A one-prompt attack that breaks LLM safety alignment

Large language models LLMs and diffusion models now power a wide range of applications, from document assistance to text-to-image generation, and users increasingly expect these systems to be safety-aligned by default. Yet safety alignment is only as robust as its weakest failure mode. Despite...

5.7AI score
SaveExploits0
Rows per page
Query Builder