arXiv · 2607.04523
Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Language
Abstract
Generic statements like "tigers are striped" and "cars have radios" communicate information that is, in general, true. However, while the first statement is true in principle, the second is true only statistically. People are exquisitely sensitive to this principled-vs-statistical distinction. It has been argued that this ability to distinguish between something being true by virtue of it being a category member versus being true because of mere statistical regularity, is a general property of people's conceptual machinery and cannot itself be learned. We investigate whether the distinction between principled and statistical properties can be learned from language itself. If so, it raises the possibility that language experience can bootstrap core conceptual distinctions and that it is possible to learn sophisticated causal models directly from language. We find that language models are all sensitive to statistical prevalence, but struggle with representing the principled-vs-statistical distinction controlling for prevalence. Until GPT-4, which succeeds.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Zhimin Hu, Jeroen van Paridon, Gary Lupyan. 2026-07-05. Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Language. https://doi.org/10.17617/2.3587960
Cite the original work for its findings. Save a collection to share your selection of sources.