cs.AISep 30, 2026

OverForge: Reasoning Through Strategies and Tactics Helps Cooperative Lifelong Adaptation

Authors: Oana Madalina Fron, Ojas Shirekar, Chirag Raman

Organizations: Tapri Lab, Department of Pattern Recognition and Bioinformatics, Delft University of Technology, Delft, The Netherlands

Abstract

Cooperative language-model agents must coordinate over long horizons and adapt to changing environments and to partners with unfamiliar conventions, yet existing agents map observations to actions without separating persistent coordination strategies from their tactical execution. We introduce OverForge, a training-free hierarchical architecture that separates strategic reasoning over roles and divisions of labour from tactical reasoning over actions within each agent's private, partner-conditioned world model. A metacognitive Prefrontal Cortex Module couples the two levels by forming strategy-action branches, imagining their consequences with a forward model, and committing when confident. In OvercookedV2, OverForge delivers 7 soups in a connected kitchen versus 3 for each flat LLM baseline, retains agreed roles, and adopts roles proposed by unfamiliar partners. Ablations and a fixed-strategy probe show that persistent strategies guide tactical adaptation while each reasoning level contributes to coordination. Memory restarts show that cross-episode partner knowledge supports task performance and partner prediction, linking the hierarchy to continual adaptation.

Figures & tables

Appendix figures & tables14 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

    Jun 6, 2026Kale-ab Abebe Tessera, Andras Szecsenyi, Cameron Barker +7Multi-Agent CoordinationMulti-Agent Large Language Model Systems

  2. SyncPlan: Long-Horizon LLM Coordination with Explicit Synchronization and Adaptive Correction

    Aug 3, 2026Shen You, Xiaoming Zhu, Weining Weng +17CoordinationGradient Staleness

  3. TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination

    May 1, 2026Yi Xie, Siao Liu, Falong Fan +3Multi-Agent Large Language Model SystemsTeaming