arXiv · 2609.16189
Position: AI Is Not Ready for Strategic Conflicts
Abstract
Open-ended strategic wargames are high-stakes LM-based social simulations: they model adversaries, institutions, escalation, plan brittleness, doctrine, and crisis response. Language models (LMs) are attractive because they can play agents, generate scenario branches, adjudicate ambiguous actions, and summarize lessons, but the same affordances make open-ended roles dangerous: model language determines both what an actor attempts and what becomes simulated reality. This position paper argues that no LM-enabled wargame should inform planning, doctrine, policy, or crisis response without an auditable safety case, and that the proper use of open-ended wargames today is to stress-test decision-influencing LM agents. We identify five failure modes: decision laundering, adjudication opacity, role collapse, escalation-through-adjudication, and failure of strategic imagination. Ordinary benchmarks cannot establish safety for these settings. Wargames can expose failures as stress tests; they are not themselves safety cases for consequential use.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Mark Riedl, Glenn Matlin. 2026-07-31. Position: AI Is Not Ready for Strategic Conflicts. https://arxiv.org/abs/2609.16189
Cite the original work for its findings. Save a collection to share your selection of sources.