physics.soc-phSep 29, 2026

Multi-agent discussion gains less when dissent is withheld

Authors: Chand Sahil Mansuri, Xin Wang, Mengying Li, Bryan Acton, Rory Eckardt, Dhaval Patel, Sadamori Kojaku

Organizations: School of Systems Science and Industrial Engineering, Binghamton University, Binghamton, NY, USA · School of Management, Binghamton University, Binghamton, NY, USA · IBM T. J. Watson Research Center, Yorktown Heights, NY, USA

Abstract

Multi-agent systems of LLMs add discussion to majority voting and are therefore expected to be more capable. However, empirical reports conflict on whether discussion improves accuracy or leads to an incorrect consensus. Here, we introduce a parsimonious model that explains when discussion improves accuracy and when it ends in an incorrect consensus, built from four behaviors repeatedly observed in LLM agents: (1) withholding dissent, (2) internalizing a stated answer, (3) reconsidering after seeing dissent, and (4) correcting toward the correct answer. The model shows that discussion can overturn an incorrect initial majority only when the withholding rate cc is below a critical rate c∗=γ/(γ+a)c^* = γ/(γ+ a), set by the net correction rate γγ and the internalization rate aa. We estimate these rates from conversation logs with a Bayesian method and place LLM teams relative to c∗c^*. As the model predicts, the gain from discussion shrinks as withholding rises, across LLMs and on a hidden profile benchmark, HiddenBench, and MedEInst. Instructing agents not to withhold dissent increases this gain. Turning reasoning off also increases the gain, because reasoning raises the internalization rate aa and keeps agents from reconsidering a minority answer. These findings reconcile the conflicting reports and identify when discussion outperforms majority voting.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Minority Sentinel: When to Overturn Majority Voting in Multi-Agent LLM Debates

    Jun 28, 2026Chuan He, Zebin Chen, Zhengyi Yang +5Multi-Agent DebateDebate

  2. Too Polite to Disagree: Understanding Sycophancy Propagation in Multi-Agent Systems

    Apr 3, 2026Vira Kasprova, Amruta Parulekar, Abdulrahman AlRabah +5SycophancyMulti-Agent Large Language Model Systems

  3. The Cost of Consensus: Isolated Self-Correction Prevails Over Unguided Homogeneous Multi-Agent Debate

    Apr 29, 2026Blaž Bertalanič, Carolina FortunaMulti-Agent DebateDebate