Lucene search
+L

1 matches found

Kitploit
Kitploit
added 2026/09/15 7:45 p.m.8 views

AutoRAN-public

🧠 AutoRAN: Automated Hijacking of Safety Reasoning in Large Reasoning Models AutoRAN is an automated Hijacking of Safety Reasoning that leverages less-aligned secondary auxiliary models to simulate reasoning traces, generate narrative prompts, and iteratively refine those prompts to bypass safety...

6AI score
SaveExploits0
Rows per page
Query Builder