Social Circuits behind Multi-agent Echo Chambers
Organizations: Simon Fraser University · University of Alberta · Southern University of Science and Technology
Abstract
Language-model agents exchange messages to combine evidence, but their communication can also create echo chambers that reinforce shared errors. However, overall task performance does not explain how a message changes the receiving agent's internal activations and affects its decision. In this work, we introduce Social Circuits, a framework for tracing message effects through receiver activations. We compare the receiver's answers before and after changing a message. Then, we restore selected activations recorded under the original message to determine how much of the message effect these activations reproduce. Based on Social Circuits, we propose Circuit-Guided Deliberation (CGD), which learns to select useful messages using receiver activation changes. We establish when activation replacement preserves receiver decisions and bound the gap between CGD's task performance and the best achievable through message selection. Experiments show that receiver activation changes explain the message effects and guide message selection that improves the task performance. Across three models and four datasets, CGD achieves the highest or joint-highest average accuracy in our main comparisons while generating fewer tokens than multi-agent baselines.
Figures & tables
| Single-agent | Multi-agent | |||||||||
| Model | Metric | Single | SC6 | Majority | MAD | MAD-M2 | MADC | MOC | SafeSieve | CGD |
| 2Wiki (A/B answers) | ||||||||||
| Gemma-3-4B | Accuracy | 78.89 | 77.78 | 83.19 | 80.69 | 82.92 | 79.44 | 86.94 | 79.72 | 89.72 |
| Tokens | 79 | 478 | 219 | 647 | 516 | 638 | 401 | 485 | 139 | |
| Qwen3-4B | Accuracy | 73.19 | 70.83 | 73.61 | 80.42 | 82.08 | 77.22 | 86.94 | 83.89 | 89.86 |
| Tokens | 171 | 1025 | 468 | 978 | 1028 | 961 | 864 | 874 | 299 | |
| Single-agent | Multi-agent | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| Model | Single | SC6 | Majority | MAD | MAD-M2 | MADC | MOC | SafeSieve | CGD |
| Gemma-3-4B | 24.72 | 25.42 | 26.39 | 7.22 | 20.97 | 7.22 | 20.28 | 25.42 | 42.92 |
| Qwen3-4B | 21.94 | 22.22 | 25.97 | 23.33 | 25.14 | 23.33 | 29.31 | 40.56 | 51.81 |
| Qwen3-8B | 23.61 | 23.89 | 30.00 | 14.17 | 30.83 | 14.17 | 28.75 | 29.86 | 61.53 |
Appendix figures & tables24 assets
Supplementary material from the paper’s appendix.
Appendix
| Symbol | Meaning |
|---|---|
| Message combination, a subset of . | |
| Combinations with the highest CGD score and expected task utility, respectively. | |
| Another message combination in . | |
| Candidate messages for one receiver. | |
| Activation-summary dimension. | |
| Incorrect answer favored by multiple agents before delivery. |
| Comparison | Task | Questions |
|---|---|---|
| Collaboration approaches | Correlated-Trap | 600 |
| Message content and repetition | Correlated-Trap | 240 per direction |
| Activation replacement | Correlated-Trap | 240 per direction |
| Message-specific replacement | 2Wiki | 240, two candidates each |
| Two-option message selection | 2Wiki | 240 per test set |
| Two-answer scoring | MATH-500 | 200 |
| Approach | Information |
|---|---|
| Single, SC6 | The designated receiver’s assigned input, without messages from other agents. |
| Majority vote | Independent answers from three agents, each using its own assigned input. |
| MAD, MADC | Each agent’s assigned input and the responses exchanged according to the communication procedure. |
| MAD-M2 | Each agent’s assigned input and the responses retained by memory masking. |
| MOC | The receiver’s assigned input, a summary of the two senders’ messages, and an updated sender response. |
| SafeSieve | Each agent’s assigned input and messages received through the learned communication graph. |
| Model | Activation changes | Answer agreement (%) | Answer recovery (%) |
|---|---|---|---|
| (a) 2Wiki | |||
| Gemma-3-4B | Original message | ||
| Another message (same task) | |||
| Qwen3-4B | Original message | ||
| Another message (same task) | |||
| (b) Correlated-Trap | |||
| Strategy | Gemma-3-4B | Qwen3-4B | Qwen3-8B |
|---|---|---|---|
| No message | 84.17 | 78.75 | 81.67 |
| Sender 1 only | 86.11 | 81.67 | 84.58 |
| Sender 2 only | 82.50 | 79.58 | 77.50 |
| Both messages | 85.28 | 83.89 | 82.78 |
| CGD | 89.72 | 91.11 | 87.36 |
| Approach | Gemma-3-4B | Qwen3-4B | Qwen3-8B |
|---|---|---|---|
| Activation products | 88.06 | 87.92 | 85.00 |
| Activation products (adjusted) | 89.31 | 88.75 | 86.94 |
| CGD | 89.72 | 89.86 | 87.36 |
| Approach | Gemma-3-4B | Qwen3-8B |
|---|---|---|
| Text-only | ||
| Receiver-only | ||
| Text + receiver | ||
| Additive | ||
| Activation products | ||
| CGD |
| Approaches | Gemma-3-4B | Qwen3-4B | Qwen3-8B |
|---|---|---|---|
| Single | 78.75 | 75.83 | 75.42 |
| MAD | 81.11 | 81.39 | 81.94 |
| MAD-M2 | 82.92 | 81.53 | 78.47 |
| MADC | 79.72 | 79.03 | 81.53 |
| MOC | 86.81 | 86.94 | 85.69 |
| SafeSieve | 80.00 | 84.17 | 81.94 |