cs.LGSep 28, 2026

Small transformers track Bayesian evidence for latent common causes via a context-invariant mechanism

Authors: Amir Mohammadpour, Michael Franke

Organizations: Department of Linguistics, University of Tübingen

Abstract

We present an in-depth investigation of how a form of Bayesian reasoning about common causes can emerge as a cross-contextual generalization in small, tractable transformers. Incrementing on recent work, our set-up (i) disentangles causal mechanisms in the model from the causal structure of the true data-generating process, (ii) orients more towards natural language prediction by considering inference of latent common causes, and (iii) considers whether and how Bayesian evidence accumulation for latent common causes can be implemented in representations and mechanisms that allow for cross-context generalization to novel test cases.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. BayesBench: Evaluating LLM Belief Trajectories Under Multi-Turn Evidence Accumulation

    Jun 29, 2026Ankur Samanta, Akshayaa Magesh, Tal Lancewicki +7LLM Reasoning Strategies

  2. For What Reason? Interpreting Models' Encoding of Causation and Antithesis

    Jul 20, 2026Abhidip Bhattacharyya, Shira WeinTransformer ArchitecturesCausal

  3. What does a Bayes-filtered transformer believe? A predictive Monte Carlo approach

    Jul 19, 2026Afiq Abdillah Effiezal Aswadi, Haotong Ma, Susan WeiPosterior Predictive DistributionTransformer Architectures