Lucene search
+L

173 matches found

Kitploit
Kitploit
added 2026/09/04 11:32 p.m.3 views

AutoRAN-public

🧠 AutoRAN: Automated Hijacking of Safety Reasoning in Large Reasoning Models AutoRAN is an automated Hijacking of Safety Reasoning that leverages less-aligned secondary auxiliary models to simulate reasoning traces, generate narrative prompts, and iteratively refine those prompts to bypass safety...

5.4AI score
SaveExploits0
Rows per page
Query Builder