cs.CLSep 23, 2026

Script Choice in LLMs: Evidence for Late-Layer Commitment

Authors: David Kletz, Sandra Mitrović, Itay Sabato, Ljiljana Dolamić, Fabio Rinaldi

Organizations: SUPSI, IDSIA, Switzerland · Independent Researcher · armasuisse, Science & Technology, Switzerland

Abstract

In this paper, we investigate how script knowledge is distributed across the layers of LLMs using two complementary interpretability methods: logistic regression probing and logit-lens analysis. Our probing experiments reveal a clear asymmetry: both the input script and the instructed output script are encoded in the earliest layers of the network, while, in contrast, commitment to the actual output script emerges only in the final layers, with the model's intermediate representations defaulting to Latin throughout most of the layers. This two-stage process is confirmed by logit-lens analyses, which show that script commitment consistently occurs at the very last layers of the LLMs. Together with the weaker script-following performance observed in smaller models, these results form a converging body of evidence linking script commitment to model depth, with broader implications for the design of sufficiently deep, inclusive multilingual architectures.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. The Latin Substrate: How Language Models Represent and Mediate Script Choice

    May 29, 2026Daniil Gurgurov, Alan Saji, Katharina Trinley +2Representational CapacityChoice

  2. Skip a Layer or Loop It? Learning Program-of-Layers in LLMs

    Jun 4, 2026Ziyue Li, Yang Li, Tianyi ZhouLLM Inference OptimizationLLM Reasoning Strategies

  3. Encoded but Not Decoded: Layer-Localized Evidence for a Three-Level Gap in LLM Syntax

    Sep 24, 2026Zhenyan Lu, He Wang, Xiaohui HuangSyntactic StructureScale Gap