Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
•added 2026/09/30 12:00 a.m.•1 views

MADBench: Benchmarking the Security of Multi-Agent Debate

Multi-agent debate MAD can improve large language model LLM reasoning by allowing multiple agents to exchange and critique their answers to the same task. However, the interactions that enable agents to correct mistakes can also spread adversarial errors and steer the agents toward an incorrect...

5.8AI score
SaveExploits0
Rows per page
Query Builder