cs.LGJan 26, 2026

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule

Authors: Yilie Huang, Wenpin Tang, Xunyu Zhou

Organizations: Department of Industrial Engineering and Operations Research, Columbia University, New York, NY 10027, USA. · Department of Industrial Engineering and Operations Research & Data Science Institute, Columbia University, New York, NY 10027, USA.

Abstract

We consider time discretization for score-based diffusion models to generate samples from a learned reverse-time dynamic on a finite grid. Uniform and hand-crafted grids can be suboptimal given a budget on the number of time steps. We introduce Adaptive Reparameterized Time (ART), which controls the clock speed of a reparameterized time variable to redistribute computation along the sampling trajectory while preserving the terminal time, with the objective of minimizing the aggregate Euler discretization error. We derive a randomized companion ART-RL that recasts ART as a continuous-time reinforcement learning problem with Gaussian policies, and prove a two-directional bridge between the two: the deterministic ART optimum lifts to an optimal Gaussian policy, and conversely any optimal Gaussian policy must recover the ART control through its mean. This bridge turns continuous-time actor--critic learning into a principled, rather than heuristic, route to the deterministic timestep optimum. Within the official EDM pipeline, ART-RL improves FID on CIFAR--10 across a wide range of budgets; after one-time offline training, the distilled deterministic schedule transfers without retraining to AFHQv2, FFHQ, and ImageNet at no extra inference cost.

Figures & tables

Explore similar work

CardsList
  1. Adaptive Reparametrized Time for Score-Based Diffusion Sampling

    Jul 2, 2026Yilie Huang, Wenpin Tang, Xun Yu ZhouDiffusion SamplingTimestep

  2. Temporal Difference Learning for Diffusion Models

    Jun 13, 2026Qizhen Ying, Yangchen Pan, Victor Adrian Prisacariu +1Denoising TrajectoryGenerative Models

  3. ReDiF: Resource-Efficient Few-Step Diffusion Distillation via Reinforcement Learning

    Dec 28, 2025Amirhossein Tighkhorshid, Zahra Dehghanian, Hamid R. RabieeDiffusion-Based Reinforcement Learning MethodsDiffusion Sampling