cs.AIOct 6, 2026

Decoupled Multi-Agent Orchestration

Authors: Xinle Wu, Yao Lu

Organizations: National University of Singapore

Abstract

Learned orchestration can automatically construct effective language-model multi-agent systems, but existing approaches couple planning to fixed worker pools and train decomposition and collaboration from the same terminal outcome, limiting transfer and obscuring credit assignment. We introduce DeOrch, which separates worker-agnostic planning from concrete worker selection. Its two-stage planner first decomposes the task without worker information, then chooses collaboration operations using compact, worker-identity-free matchability feedback from the pool, enabling conditional credit assignment to decomposition and collaboration decisions. A lightweight matcher estimates worker suitability from behavior on a fixed probe set and adapts online with a contextual bandit, allowing new workers to be incorporated without retraining the planner or matcher. Across diverse in- and out-of-distribution tasks, DeOrch outperforms prior automatic MAS orchestration methods with fewer worker calls than competing learned orchestrators, remains effective when transferred to an entirely unseen worker pool without retraining, and shows consistent gains from both components.

Figures & tables

Appendix figures & tables5 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Reward Modeling for Multi-Agent Orchestration

    Jun 11, 2026King Yeung Tsang, Zihao Zhao, Vishal Venkataramani +5Multi-Agent Orchestration

  2. Small Model as Master Orchestrator: Learning Unified Agent-Tool Orchestration with Parallel Subtask Decomposition

    Apr 18, 2026Wenzhen Yuan, Wutao Xiong, Fanchen Yu +7Multi-Agent OrchestrationOrchestration