arXiv · 2609.39727
OverForge: Reasoning Through Strategies and Tactics Helps Cooperative Lifelong Adaptation
Abstract
Cooperative language-model agents must coordinate over long horizons and adapt to changing environments and to partners with unfamiliar conventions, yet existing agents map observations to actions without separating persistent coordination strategies from their tactical execution. We introduce OverForge, a training-free hierarchical architecture that separates strategic reasoning over roles and divisions of labour from tactical reasoning over actions within each agent's private, partner-conditioned world model. A metacognitive Prefrontal Cortex Module couples the two levels by forming strategy-action branches, imagining their consequences with a forward model, and committing when confident. In OvercookedV2, OverForge delivers 7 soups in a connected kitchen versus 3 for each flat LLM baseline, retains agreed roles, and adopts roles proposed by unfamiliar partners. Ablations and a fixed-strategy probe show that persistent strategies guide tactical adaptation while each reasoning level contributes to coordination. Memory restarts show that cross-episode partner knowledge supports task performance and partner prediction, linking the hierarchy to continual adaptation.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Oana Madalina Fron, Ojas Shirekar, Chirag Raman. 2026-09-30. OverForge: Reasoning Through Strategies and Tactics Helps Cooperative Lifelong Adaptation. https://arxiv.org/abs/2609.39727
Cite the original work for its findings. Save a collection to share your selection of sources.