cs.CLOct 1, 2026

Know When to Hold 'em: Correct-Token Retention in Uniform-State Diffusion Language Models

Authors: Mojtaba Nafez, James Henderson

Organizations: EPFL & Idiap Lausanne, Switzerland · Idiap Research Institute Martigny, Switzerland

Abstract

Uniform-state diffusion models (USDMs) can revise any token at any denoising step, which lets them correct their own mistakes, a key advantage over masked diffusion. Self-correction, however, requires both revising incorrect tokens and retaining correct ones, and we show that current USDMs lack the latter. Even under greedy-tail decoding, state-of-the-art USDMs (DUO, UDLM, and uniform-noise SEDD) keep revising 173--270 of 512 positions at every step, and these large, uncoordinated edits collapse sample diversity. A random-token corruption experiment traces this deficit to the models themselves: they reconstruct clean and corrupted tokens with nearly identical accuracy, even though clean tokens are easier targets. A decomposition of the validation NELBO shows that training barely rewards retention: incorrect predictions are heavily penalized at corrupted positions but almost free at clean ones. We propose Correct-Token Retention Regularization (CTR-Reg), a simple but effective auxiliary loss that trains the model to retain tokens left unperturbed by the forward process and requires no change to the sampler. CTR-Reg improves clean-token accuracy by 26.5 percentage points on average across six benchmarks, while leaving corrupted-token accuracy virtually unchanged, and its per-step revisions converge to only 3--11 positions. With just five greedy-tail steps, generative perplexity more than halves under CTR-Reg for all three models while diversity is preserved, and these gains hold across sampling budgets. Our results identify correct-token retention as a key missing ingredient for self-correcting diffusion language models, and demonstrate an effective fix.

Figures & tables

Appendix figures & tables22 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Remask, Don't Replace: Token-to-Mask Refinement in Diffusion Large Language Models

    Apr 20, 2026Lin YaoDiffusion Language ModelsIterative Self-Correction

  2. BlockGen: Flexible Blockwise Sequence Modeling with Hybrid Samplers

    Jun 1, 2026Justin Deschenaux, Caglar GulcehreDiffusion ModelsDiffusion Language Models