Lucene search
+L

2 matches found

Kitploit
Kitploit
•added 2026/09/28 11:10 p.m.•11 views

gradient-untangler

gradient-untangler A local research harness that searches for the exact tokens that make an open-weight language model start its answer the way you specify. Open-weight models ship as ordinary files: config.json, a tokenizer, and one or more .safetensors or .bin shards. A Hugging Face id is only ...

6.2AI score
SaveExploits0References7
Packet Storm News
Packet Storm News
•added 2025/11/23 12:00 a.m.•39 views

TASO: Jailbreak LLMs Via Alternative Template and Suffix Optimization

Many recent studies showed that LLMs are vulnerable to jailbreak attacks, where an attacker can perturb the input of an LLM to induce it to generate an output for a harmful question. In general, existing jailbreak techniques either optimize a semantic template intended to induce the LLM to produc...

7AI score
SaveExploits0
Rows per page
Query Builder