














Abstract:Large Language Models (LLMs) have advanced autonomous agents' planning and decision-making, yet they struggle with complex tasks requiring diverse expertise and multi-step reasoning. Multi-Agent Debate (MAD) systems, introduced in NLP research, address this gap by enabling structured debates among LLM-based agents to refine solutions iteratively. MAD promotes divergent thinking through role-specific agents, dynamic interactions, and structured decision-making. Recognizing parallels between Software Engineering (SE) and collaborative human problem-solving, this study investigates MAD's effectiveness on four coding tasks in SE. We adapt a MAD framework from NLP, analyze agent interactions to assess consensus-building and iterative refinement, and propose two MAD variants that enhance agent debate for coding tasks by addressing the observed weaknesses. Our findings show that structured debate and collaboration improve problem-solving and yield strong performance in some cases, highlighting the collaborative debate synergy between LLM agents for coding tasks in SE while identifying areas for future exploration.
From: Yong Jin (Jina) Chun [view email]
[v1]
Sat, 15 Mar 2025 07:30:37 UTC (538 KB)
[v2]
Mon, 3 Aug 2026 22:16:33 UTC (207 KB)
此内容由惯性聚合(RSS阅读器)自动聚合整理,仅供阅读参考。 原文来自 — 版权归原作者所有。