Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2026/05/18 12:0 a.m.21 views

Be Kind, Rewrite: Benign Projections Via Rewriting Defend against LLM Data Poisoning Attacks

Large language models LLMs are highly susceptible to backdoor attacks BAs, wherein training samples are poisoned using trigger-based harmful content. Furthermore, existing defenses have proven ineffective when extensively tested across BA patterns. To better combat BAs, we explore the use of LLM...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/07/06 12:0 a.m.9 views

UniAud: a Unified Auditing Framework for High Auditing Power and Utility with One Training Run

Differentially private DP optimization has been widely adopted as a standard approach to provide rigorous privacy guarantees for training datasets. DP auditing verifies whether a model trained with DP optimization satisfies its claimed privacy level by estimating empirical privacy lower bounds...

6.8AI score
SaveExploits0
Rows per page
Query Builder