arXiv · 2503.02390
ReSo: A Reward-driven Self-organizing LLM-based Multi-Agent System for Reasoning Tasks
Abstract
Multi-agent systems (MAS) have emerged as a promising approach for enhancing the reasoning capabilities of large language models in complex problem-solving; however, current MAS frameworks suffer from poor flexibility and scalability with underdeveloped optimization strategies. To address these challenges, we propose ReSo, which integrates task graph generation with a reward-driven two-stage agent selection process centered on our Collaborative Reward Model that provides fine-grained reward signals to optimize MAS cooperation. We also introduce an automated data synthesis framework for generating MAS benchmarks without any human annotations. Experimental results show that ReSo matches or outperforms existing methods, achieving 33.7 percent accuracy on Math-MAS and 32.3 percent accuracy on SciBench-MAS, where other approaches completely fail.
Explore related subjects
Keep this discovery
Heng Zhou, Hejia Geng, Xiangyuan Xue, Li Kang, Yiran Qin, Zhiyong Wang, Zhenfei Yin, Lei Bai. 2025-03-04. ReSo: A Reward-driven Self-organizing LLM-based Multi-Agent System for Reasoning Tasks. https://arxiv.org/abs/2503.02390
Cite the original work for its findings. Save a collection to share your selection of sources.