cs.AIAug 12, 2025

Reducing Cognitive Overhead in Tool Use via Multi-Small-Agent Reinforcement Learning

Authors: Dayu Wang, Yutong Liu, Jiaye Yang, Weikang Li, Jiahui Liang, Yang Li

Organizations: Baidu Inc. · Peking University

Abstract

Recent advances in multi-agent systems highlight the potential of specialized small agents that collaborate via division of labor. Existing tool-integrated reasoning systems, however, often follow a single-agent paradigm in which one large model interleaves long-horizon reasoning with precise tool operations, leading to cognitive-load interference and unstable coordination. We present MSARL, a Multi-Small-Agent Reinforcement Learning framework that explicitly decouples reasoning from tool use. In MSARL, a Reasoning Agent decomposes problems and plans tool invocations, while multiple Tool Agents specialize in specific external tools, each trained via a combination of imitation learning and reinforcement learning with role-specific rewards. On mathematical problem solving with code execution, MSARL significantly improves reasoning stability and final-answer accuracy over single-agent baselines. Moreover, the architecture generalizes to diverse tool-use tasks, demonstrating that cognitive-role decoupling with small agents is a scalable blueprint for multi-agent AI design.

Figures & tables

Explore similar work

CardsList
  1. Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning

    Feb 1, 2026Yu Li, Mingyang Yi, Xiuyu Li +6Agentic Reinforcement LearningAgentic Reasoning

  2. Token-Efficient Multi-Agent Collaboration via System One-Guided Computational Division of Labor

    Oct 6, 2026Zihan Zhou, Xinzhe Hu, Hanxu Yang +2Multi-Agent Large Language Model SystemsModel-Based Multi-Agent Systems

  3. Benefits and Limitations of Communication in Multi-Agent Reasoning

    Oct 14, 2025Michael Rizvi-Martel, Satwik Bhattamishra, Neil Rathi +2Multi-Agent ReasoningMulti-Hop Reasoning