Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
•added 2026/10/05 12:00 a.m.•6 views

RAISED: Self-Distillation for Robustness to Prompt Injection in LLM Agents

Tool-using language-model agents are vulnerable to indirect prompt injection because they must act on untrusted external content. Existing training-time defenses can reduce attack success rates, but often at the cost of general capabilities. We show that training-based defenses induce substantial...

5.9AI score
SaveExploits0
Rows per page
Query Builder