arXiv · 2609.35351
Interference Beyond Geometry in Concept Extraction
Abstract
Interference is commonly treated as geometric overlap between learned features. We introduce effective interference, which combines feature geometry and code statistics to capture realized interactions, distinguishing constructive from destructive interference and frequent weak interactions from rare strong ones. Under local fixed-support assumptions, we characterize how architectural constraints shape interference through four mechanisms: feature orthogonalization, bias compensation, gain adaptation, and encoder-decoder separation. Experiments with sparse autoencoders show that constrained architectures selectively reduce overlap among co-active features, while bias, gain, and encoder freedom allow constructive cross-contributions to remain. Together, these results show that interference in learned representations depends not only on feature geometry, but also on how features are used and on the architecture that produces their codes.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Valérie Costa, Bahareh Tolooshams. 2026-09-28. Interference Beyond Geometry in Concept Extraction. https://arxiv.org/abs/2609.35351
Cite the original work for its findings. Save a collection to share your selection of sources.