Lucene search
+L

2 matches found

Kitploit
Kitploit
added 2026/09/07 11:09 p.m.4 views

gradient-untangler

gradient-untangler A local research harness that searches for the exact tokens that make an open-weight language model start its answer the way you specify. Open-weight models ship as ordinary files: config.json, a tokenizer, and one or more .safetensors or .bin shards. A Hugging Face id is only ...

6AI score
SaveExploits0References10
Packet Storm News
Packet Storm News
added 2025/11/23 12:00 a.m.37 views

TASO: Jailbreak LLMs Via Alternative Template and Suffix Optimization

Many recent studies showed that LLMs are vulnerable to jailbreak attacks, where an attacker can perturb the input of an LLM to induce it to generate an output for a harmful question. In general, existing jailbreak techniques either optimize a semantic template intended to induce the LLM to produc...

7AI score
SaveExploits0
Rows per page
Query Builder