arXiv · 2607.00233
From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents
Abstract
How do two agents invent a shared language from scratch? In a Lewis signaling game, a sender and receiver must coordinate on a code using only their interaction history. We study five memory architectures across varying channel configurations with LLM agents and find that memory architecture matters more than channel capacity. Agents with a persistent private notebook benefit from surplus channel capacity and avoid the high-capacity collapse seen in stateless agents, achieving the most reliable coordination ($0.867 \pm 0.023$ at capacity = 25). Stateless agents peak at moderate capacity and then degrade as the vocabulary grows beyond what a rolling context window can track The notebook externalizes learned conventions, freeing agents from having to re-derive codes each round. An information bottleneck-inspired argument predicts an optimal capacity equal to the number of objects. Instead, the bottleneck (capacity = 8) proves to be a fragility point, and surplus capacity is generally better. We show that channel capacity alone cannot predict coordination; memory architecture determines whether agents turn interaction history into stable conventions, and both dimensions are needed to understand how signals become language.
Explore related subjects
Keep this discovery
Yashar Talebirad, Eden Redman, Ali Parsaee, Osmar R. Zaiane. 2026-06-30. From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents. https://arxiv.org/abs/2607.00233
Cite the original work for its findings. Save a collection to share your selection of sources.