2 matches found
prompt_injection
Cross-Surface Prompt Injection Benchmark Benchmark for measuring where in a tool-using LLM agent's execution pipeline a prompt injection defense fires — and where it doesn't. Embeds unique canary tokens SECRET-A-F0-98 in injected payloads and tracks them at four pipeline stages: exposed → persist...
6.3AI score
SaveExploits0
MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG
Multimodal agentic retrieval-augmented generation RAG systems expand the attack surface beyond prompt injection to include text poisoning, image injection, direct-query attacks, and orchestrator-level tool manipulation. Existing red-teaming approaches are typically surface-specific and often...
5.8AI score
SaveExploits0
20