Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2025/08/03 12:0 a.m.8 views

IMU: Influence-Guided Machine Unlearning

Recent studies have shown that deep learning models are vulnerable to attacks and tend to memorize training data points, raising significant concerns about privacy leakage. This motivates the development of machine unlearning MU, i.e., a paradigm that enables models to selectively forget specific...

6.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/05/17 12:0 a.m.10 views

Self-Destructive Language Model

Harmful fine-tuning attacks pose a major threat to the security of large language models LLMs, allowing adversaries to compromise safety guardrails with minimal harmful data. While existing defenses attempt to reinforce LLM alignment, they fail to address models' inherent "trainability" on harmfu...

7.3AI score
SaveExploits0
Rows per page
Query Builder