Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/05/25 12:0 a.m.8 views

ALRPHFS: Adversarially Learned Risk Patterns with Hierarchical Fast \& Slow Reasoning for Robust Agent Defense

LLM Agents are becoming central to intelligent systems. However, their deployment raises serious safety concerns. Existing defenses largely rely on "Safety Checks", which struggle to capture the complex semantic risks posed by harmful user inputs or unsafe agent behaviors - creating a significant...

7.2AI score
SaveExploits0
Rows per page
Query Builder