cs.CLJul 22, 2026

surprisal is Not a Theory

Authors: Andrés Buxó-LugoAniello De SantoMorgan GrobolRyan J. HubbardCassandra L. Jacobs

Organizations: aUniversity at Buffalo, Buffalo, NY, USA · bUniversity of Utah, Salt Lake City, UT, USA · cUniversité Paris Nanterre, Paris, France · dUniversity at Albany, Albany, NY, USA

Abstract

Surprisal Theory is often characterized as a computational-level explanation per (Marr, 1982). We argue in this work that, even though a computational level narrative has been used to support "representation-agnostic research" within computational psycholinguistics, the movement toward black box systems embodied by large language models (LLMs) does not exempt modelers using the surprisal metric from the representational decisions required by computational-level characterizations. In fact, we argue that the uncritical use of LLM-surprisal obfuscates the representational and algorithmic-level commitments of different models. In three analyses, we show that the choice of algorithm and model architecture play significant roles in the computation of language model probabilities. We advise that researchers who wish to test Surprisal Theory re-evaluate the practice of treating large language model probabilities as interchangeable

Explore similar work

CardsList
  1. On the Proper Treatment of Units in Surprisal Theory

    Apr 30, 2026Samuel Kiegeland, Vésteinn Snæbjarnarson, Tim Vieira +1SurprisalTheoretical Foundations