Lucene search
+L

171 matches found

Kitploit
Kitploit
added 2026/09/03 2:22 a.m.3 views

AutoRAN-public

🧠 AutoRAN: Automated Hijacking of Safety Reasoning in Large Reasoning Models AutoRAN is an automated Hijacking of Safety Reasoning that leverages less-aligned secondary auxiliary models to simulate reasoning traces, generate narrative prompts, and iteratively refine those prompts to bypass safety...

5.9AI score
SaveExploits0
Rows per page
Query Builder