cs.CLOct 1, 2026

Temporally-Resolved Token Attribution Reveals the Generation Dynamics of Diffusion Language Models

Authors: Darpan Aswal, Céline Hudelot

Organizations: Université Grenoble Alpes · MICS, CentraleSupélec

Abstract

This work presents Diffusion Layer Integrated Gradients (DLIG), a token attribution method for diffusion language models (DLMs) that extends Integrated Gradients (IG~\cite{sundararajan2017axiomatic}) to arbitrary layers and denoising steps. DLIG attributes a DLM's progressive commitment to a self-generated or fixed completion for an input prompt. We establish direct correspondences between DLIG and the IG axioms of completeness, implementation invariance, linearity, and symmetry preservation. As a lightweight complement to interventional analysis, DLIG provides an inexpensive first check of mechanistic hypotheses across the denoising trajectory. We demonstrate this on word-sense disambiguation, multi-hop graph reasoning, and sentence infilling, revealing how DLMs draw on inputs across positions, layers, and denoising steps.

Figures & tables

Appendix figures & tables19 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. DA-DLM: Explicitly Modeling Token Dependencies in Diffusion Language Models

    Sep 14, 2026Pengyu Ji, Zichen Zhang, Xiang Hu +1Diffusion Language ModelsAutoregressive Transformers

  2. Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

    Jun 10, 2026Jia Deng, Junyi Li, Wayne Xin Zhao +3Diffusion Language Models

  3. Measuring Temporal Linguistic Emergence in Diffusion Language Models

    Apr 25, 2026Harry LuDiffusion Language ModelsLarge Language Model Uncertainty