Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2026/02/16 12:0 a.m.6 views

Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks

As the capabilities of large language models continue to advance, so does their potential for misuse. While closed-source models typically rely on external defenses, open-weight models must primarily depend on internal safeguards to mitigate harmful behavior. Prior red-teaming research has largel...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/22 12:0 a.m.5 views

When Forgetting Triggers Backdoors: a Clean Unlearning Attack

Machine unlearning has emerged as a key component in ensuring Right to be Forgotten, enabling the removal of specific data points from trained models. However, even when the unlearning is performed without poisoning the forget-set clean unlearning, it can be exploited for stealthy attacks that...

7AI score
SaveExploits0
Rows per page
Query Builder