arXiv · 2605.22532
Relational Linear Properties in Language Models: An Empirical Investigation
Abstract
Linear properties are ubiquitous in the representations of language models; however, testing them experimentally remains a challenging task. This work focuses on relational linearity: the hypothesis that, for a fixed relation (e.g., "plays"), the unembedding of an object (e.g., "trumpet") can be predicted from the embedding of its subject (e.g.,"Miles Davis") by a linear map. We present an experimental method to test the formulation of relational linearity by Marconato et al. (2025). Specifically, we introduce a probing method, based on Kullback-Leibler divergence, to evaluate this property and examine its variation across layers and paraphrased relational queries. It is also more efficient than previous work; for example, it avoids the crude Jacobian approximations used in Linear Relational Embeddings by Hernandez et al. (2024). Our findings across four datasets show that relational linearity varies across models, exhibits layer-wise patterns consistent with prior observations about linguistic information in model representations, and is differently affected by changes in how the relation is phrased.
Explore related subjects
Keep this discovery
Giovanni Valer, Luigi Gresele, Marco Bronzini, Emanuele Marconato. 2026-05-21. Relational Linear Properties in Language Models: An Empirical Investigation. https://arxiv.org/abs/2605.22532
Cite the original work for its findings. Save a collection to share your selection of sources.