arXiv · 2609.20831
Recursive Language Models Generalize Out of Domain
Abstract
We study when limiting what a language model can see improves learning. We compare standard CoT, the more general learner that reads the full trace, with recursive language models, which restricts itself by solving each subtask in an isolated context. In-distribution, this generality comes for free: CoT can efficiently simulate the recursive rule, so the IID generalization guarantee changes only by a constant factor, and recursion does not offer much. But out of domain, CoT can fit training by relying on context outside the current subtask, i.e. a shortcut that breaks once those tokens change; recursive context isolation rules out this failure mode. Even though CoT's class still covers the recursive rule, simplicity bias picks the shortcut over the truth. Thus, to go beyond distributional accuracy and truly reason, covering the right rule is not enough; this contrasts with classical learning theory.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Chenxiao Yang, Zhiyuan Li, David McAllester, Nathan Srebro. 2026-07-23. Recursive Language Models Generalize Out of Domain. https://arxiv.org/abs/2609.20831
Cite the original work for its findings. Save a collection to share your selection of sources.