cs.CLOct 6, 2026

STRUCTURALCOST: A controlled reading time dataset for modeling human sentence processing difficulty

Authors: Nina Nusbaumer, Iria de-Dios-Flores, Corentin Bel, Christophe Pallier, Guillaume Wisniewski, Benoît Crabbé

Organizations: LLF, CNRS, Université Paris Cité · COLT, Universitat Pompeu Fabra · Unicog, Neurospin, CEA · LNC2, ENS-PSL · INSERM, CNRS

Abstract

We introduce STRUCTURALCOST, a self-paced reading dataset of 475 participants and 40,800 observations isolating the processing cost of long-distance subject-verb dependency resolution. We replicate a low-powered psycholinguistic finding at NLP scale, namely that human reading times at the main verb increase with dependency length, driven by syntactic embedding beyond linear distance. Different language models -- spanning n-gram models, SSMs, and transformers -- partially mirror this graded difficulty profile, yet underestimate the integration cost humans incur, with a gap that persists across architectures and model sizes. This suggests these models capture the predictive component of human processing but not the full integration cost that working memory imposes. STRUCTURALCOST provides data needed to drive progress toward evaluating the cognitive plausibility of language models.

Figures & tables

Appendix figures & tables3 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Dual Alignment Between Language Model Layers and Human Sentence Processing

    Apr 20, 2026Tatsuki Kuribayashi, Alex Warstadt, Yohei Oseki +1Natural LanguageSurprisal

  2. Probing for Reading Times

    Apr 20, 2026Eleftheria Tsipidi, Samuel Kiegeland, Francesco Ignazio Re +4ReadabilityEye Tracking

  3. Trajectory Dynamics in Language Model Hidden States Predict Human Processing Costs Beyond Surprisal

    Jun 3, 2026Elan BarenholtzSurprisalLanguage Modeling