arXiv · 2610.02040
Typological Alignment of Stack-Based Language Models on Mildly Context-Sensitive Artificial Languages
Abstract
Some properties of languages, e.g., subject-object-verb (SOV) word order, are more prevalent than others among the thousands of attested natural languages (NLs). Such typological commonality is often attributed to learning biases. Computational simulations, recently with language models (LMs), have facilitated the exploration of this theory. In this paper, we extend existing analyses of the relationship between LMs' learning biases and typological commonality on both data and model sides, focusing on: (i) cross-serial dependencies, the upper limit of attested syntactic complexity, and (ii) stack-based LMs (SLMs), potentially facilitating learning of hierarchical patterns. We first evaluate generalization of SLMs on cross-serial dependencies across diverse artificial languages and confirm that they struggle with such constructions. However, SLMs with limited working memory generalize better suggesting a possible basis for such inductive bias and thus the typological commonality of some word order configurations.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Nadine El-Naggar, Tatsuki Kuribayashi, Ted Briscoe. 2026-10-01. Typological Alignment of Stack-Based Language Models on Mildly Context-Sensitive Artificial Languages. https://arxiv.org/abs/2610.02040
Cite the original work for its findings. Save a collection to share your selection of sources.