1 matches found
MADBench: Benchmarking the Security of Multi-Agent Debate
Multi-agent debate MAD can improve large language model LLM reasoning by allowing multiple agents to exchange and critique their answers to the same task. However, the interactions that enable agents to correct mistakes can also spread adversarial errors and steer the agents toward an incorrect...
5.8AI score
SaveExploits0
20