Lucene search
+L

2 matches found

Kitploit
Kitploit
added 2026/09/10 11:23 a.m.4 views

AgentWatcher

AgentWatcher AgentWatcher is a detection-based defense against indirect prompt injection in LLM agents. It first runs causal context attribution over untrusted context to find the most influential contexts, then applies a monitor LLM that classifies those contexts under explicit, customizable...

6AI score
SaveExploits0References5
Packet Storm News
Packet Storm News
added 2026/03/11 12:00 a.m.21 views

AttriGuard: Defeating Indirect Prompt Injection in LLM Agents Via Causal Attribution of Tool Invocations

LLM agents are highly vulnerable to Indirect Prompt Injection IPI, where adversaries embed malicious directives in untrusted tool outputs to hijack execution. Most existing defenses treat IPI as an input-level semantic discrimination problem, which often fails to generalize to unseen payloads. We...

5.8AI score
SaveExploits0
Rows per page
Query Builder