Lucene search
+L

2 matches found

Kitploit
Kitploit
added 2026/09/20 1:24 a.m.8 views

AgentWatcher

AgentWatcher AgentWatcher is a detection-based defense against indirect prompt injection in LLM agents. It first runs causal context attribution over untrusted context to find the most influential contexts, then applies a monitor LLM that classifies those contexts under explicit, customizable...

6.1AI score
SaveExploits0References5
Packet Storm News
Packet Storm News
added 2026/04/02 12:00 a.m.13 views

AgentWatcher: A Rule-Based Prompt Injection Monitor

Large language models LLMs and their applications, such as agents, are highly vulnerable to prompt injection attacks. State-of-the-art prompt injection detection methods have the following limitations: 1 their effectiveness degrades significantly as context length increases, and 2 they lack...

5.9AI score
SaveExploits0
Rows per page
Query Builder