arXiv · 2609.36451
Invariant Atoms: Sparse Coordinates of Local Semantic Geometry in Language Model Representations
Abstract
Large language models often preserve meaning despite substantial changes in wording, style, and syntax, while small semantic edits can systematically alter their hidden representations. This suggests that semantic variation may be organized along recurring local directions. We propose the Invariant Atom Hypothesis: local semantic motion admits preferred sparse coordinates along directions that remain stable under meaning-preserving transformations. We learn a shared semantic frame and sparse coordinates that reconstruct semantic displacements while suppressing nuisance variation, with anchor-dependent diagonal modulation adjusting atom strengths without sample-specific rotations. Empirically, the atoms exhibit strong semantic--nuisance separation, sparse reconstruction, reproducible directions, and causal effects on model predictions. The learned geometry generalizes to unseen semantic neighborhoods and nuisance families, while local reweighting improves semantic selectivity and preserves a consistent global-to-local structure. Atom signatures also remain stable under model modification. These findings support reusable invariant directions as a sparse coordinate system for local semantic geometry in language models.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Muhammad Ahtesham, Xin Zhong. 2026-09-29. Invariant Atoms: Sparse Coordinates of Local Semantic Geometry in Language Model Representations. https://arxiv.org/abs/2609.36451
Cite the original work for its findings. Save a collection to share your selection of sources.