cs.CLOct 1, 2026

Typological Alignment of Stack-Based Language Models on Mildly Context-Sensitive Artificial Languages

Authors: Nadine El-Naggar, Tatsuki Kuribayashi, Ted Briscoe

Organizations: Mohamed bin Zayed University of Artificial Intelligence · Tohoku University

Abstract

Some properties of languages, e.g., subject-object-verb (SOV) word order, are more prevalent than others among the thousands of attested natural languages (NLs). Such typological commonality is often attributed to learning biases. Computational simulations, recently with language models (LMs), have facilitated the exploration of this theory. In this paper, we extend existing analyses of the relationship between LMs' learning biases and typological commonality on both data and model sides, focusing on: (i) cross-serial dependencies, the upper limit of attested syntactic complexity, and (ii) stack-based LMs (SLMs), potentially facilitating learning of hierarchical patterns. We first evaluate generalization of SLMs on cross-serial dependencies across diverse artificial languages and confirm that they struggle with such constructions. However, SLMs with limited working memory generalize better suggesting a possible basis for such inductive bias and thus the typological commonality of some word order configurations.

Figures & tables

Appendix figures & tables5 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. What Kind of Language is Easy to Language-Model Under Curriculum Learning?

    Apr 29, 2026Nadine El-Naggar, Tatsuki Kuribayashi, Ted BriscoeLanguage AcquisitionTypology

  2. Left-Branching Transformers Excel at Right-Branching Languages: Data Shapes Word Order Preferences in Language Models

    Date pendingVarvara Arzt, Allan Hanbury, Terra BlevinsLanguage ModelingLanguage Pairs

  3. Language Models Generalize to Human-like Word Order Preferences

    Aug 5, 2026Amanda Popadich, Shane Steinert-ThrelkeldLanguage ModelingLarge Language Model Bias