arXiv · 2609.33974
Beyond Solo and Consistency: Vindicating Multi-Agent Debate via Conditional Progressive Pruning
Abstract
Large Language Model (LLM) based Multi-Agent Debate (MAD) is one of the most effective test time scaling techniques. Through multi-round communication, agents complement each other in knowledge and reasoning and solve tasks that no single member can solve. However, existing MAD frameworks fail to beat strong Single Agent and Consistency-based baselines under the same strict cost limit, which shakes the foundation of the MAD field. We propose Conditional Progressive Pruning (CPP), a lightweight pruning framework that fully exploits multi-round MAD. CPP outperforms all existing MAD frameworks on multiple dominated benchmarks. It is also the first to fully outperform consistency methods. Our code, detailed agent interaction records will be released soon.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ruosong Ye, Caiqi Zhang, Jiahao Li, Haijun Wu, Xiaolong Luo, Huiyuan Chen, Yu Wang, Ying Chen, Zhenting Wang, Kai Mei, Yang Zhou, Dimitris N. Metaxas. 2026-09-27. Beyond Solo and Consistency: Vindicating Multi-Agent Debate via Conditional Progressive Pruning. https://arxiv.org/abs/2609.33974
Cite the original work for its findings. Save a collection to share your selection of sources.