← 返回论文检索
ICLR 2025Blog Track PosterAccept

Multi-LLM-Agents Debate - Performance, Efficiency, and Scaling Challenges

Hangfan Zhang, Zhiyao Cui, Qiaosheng Zhang, Shuyue Hu

Pennsylvania State University · Northwestern Polytechnical University · Shanghai Artificial Intelligence Laboratory

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

Multi-Agent Debate (MAD) explores leveraging collaboration among multiple large language model (LLM) agents to improve test-time performance without additional training. This blog evaluates five MAD frameworks across nine benchmarks, revealing that current MAD methods fail to consistently outperform simpler single-agent strategies, even with increased computational resources. Analysis of factors such as agent configurations and debate rounds suggests that existing MAD designs fall short in fully utilizing additional inference-time computation.

论文信息

会议
ICLR 2025
年份
2025