cs.LGOct 4, 2026

LiFT: Loop Flow Transformers

Authors: Mohammad Mahdi Derakhshani, Pedro M. P. Curvo, Gertjan J. Burghouts, Jan-Willem van de Meent, Cees G. M. Snoek

Organizations: VISLab · University of Amsterdam · TNO, Intelligent Imaging · AMLab

Abstract

We introduce Loop Flow Transformers (LiFT), a family of looped generative models that scales computation by repeatedly applying a shared Diffusion Transformer (DiT) core, with only light changes to the standard architecture. Rather than asking every recurrent step for the final prediction, LiFT trains each step with a single regression target: a point on a straight path from the model's initial estimate to the flow-matching target. Because we index these targets by a continuous depth coordinate, a trained model can loop far beyond its training depth with no retraining, early exits, or other modifications. In our experiments, these longer rollouts improve generation, so inference computation can grow without adding parameters. On ImageNet at 256x256, LiFT-L/2 achieves an FID 3.34 points lower than our dense DiT-XL/2 baseline while using approximately 60% fewer parameters, 32% fewer training FLOPs, and 52% fewer inference FLOPs.

Figures & tables

Appendix figures & tables22 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Looped Diffusion Transformer

    Sep 30, 2026Yong Xien Chng, Tianyi Chen, Wenwen Tong +7Diffusion Transformers

  2. Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models

    Apr 29, 2026Zhirong Shen, Rui Huang, Jiacheng Liu +6Diffusion TransformersDiffusion Models

  3. FlashLoop: Fast and Memory-Efficient Looped Transformers via Lazy Updates

    Sep 24, 2026Wanqi Yang, Shiwei LiuBlock Sparse Flash AttentionTransformer Architectures