1 matches found
No One Architecture Fits All: A Cross-Environment Evaluation of Hierarchical Red Team Agents
Autonomous red team agents increasingly stress-test AI-enabled cyber defenses by planning strategy and executing multistage attacks. Reinforcement learning RL and large language models LLMs offer complementary mechanisms for the planning and execution such agents require, and prior work has...
5.9AI score
SaveExploits0
20