cs.CVSep 29, 2026

SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video

Authors: Haozhe Liu, Tian Ye, Shuchen Xue, Yitong Li, Junsong Chen, Haopeng Li, Jincheng Yu, Duomin Wang, +4 more

Organizations: NVIDIA

Abstract

High-resolution video generation is expensive, as its cost grows rapidly with the number of spatiotemporal tokens. A practical alternative first generates a lower-resolution video and then applies a refiner, but conventional multi-step refinement introduces a second sampling bottleneck. We present SoL-Refiner, a one-step video refiner that transforms low-resolution model outputs into 4K videos with a single denoising step. Our three-stage recipe combines high-resolution continual training, reinforcement learning (RL) post-training, and a final one-step distillation. We introduce Refiner-Bench, a video refinement benchmark constructed from the outputs of different video generators, and use a shared-input protocol to compare refiners at approximately 2K output resolution. At 2K, the one-step SoL-Refiner outperforms all external refiners on the VBench and UniPercept averages, while at 3840 ⁣× ⁣21763840\!\times\!2176 it improves both metrics over the three-step LTX-2.3 Refiner. With the complete acceleration stack, SoL-Refiner achieves an 8.91×8.91\times speedup in refinement latency over the same baseline in our 2K latency setting.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution

    May 25, 2026Wenxue Li, Jingjing Ren, Peng Zhang +4Video GenerationVisual Fidelity

  2. Dynamic Video Generation: Shaping Video Generation Across Time and Space

    May 20, 2026Shikang Zheng, Jingkai Huang, Jiacheng Liu +5Video Generation

  3. Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

    Jun 21, 2026Yitong Li, Junsong Chen, Haopeng Li +6Video Diffusion Models