cs.CVOct 5, 2026

CentriQ: Calibration-Free Quantization of Diffusion Transformers via Exact Mean Centering

Authors: Nataša Jovanović, Mathieu Salzmann, Saqib Javed

Organizations: Tenstorrent · EPFL

Abstract

Diffusion transformers (DiTs) achieve state-of-the-art image generation, but their sampling cost limits deployment. Quantizing both weights and activations to 4 bits reduces this cost, yet existing methods fall short in one of two ways. Calibration-based methods are tied to a specific checkpoint and prompt distribution, whereas data-free Hadamard rotation, effective for LLMs, loses quality on DiTs. We show that this loss has a structural cause. Adaptive layer-norm conditioning adds a per-token mean to the activations, and at the widths of the evaluated DiTs, the Hadamard rotations used by data-free methods cannot spread this mean uniformly across coordinates. A single dominant direction therefore survives the rotation and sets the quantization range. We introduce CentriQ, a calibration-free quantizer that centers each token before rotation and restores the mean exactly through a rank-1 full-precision branch, so that per-token scales follow in closed form without data. Weights are fitted under a robust ℓp\ell_p objective that tracks the dense mode of each group and discounts heavy tails. Across three DiTs, CentriQ matches the quality of calibrated SVDQuant at 4 bits, whereas calibration-free weight quantizers with plain per-token activation quantization collapse or degrade substantially. CentriQ outperforms the strongest calibration-free method reported to date at 2-bit weights. It is also the first calibration-free method to retain usable image quality at 2-bit activations.

Figures & tables

Appendix figures & tables15 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers

    Jul 2, 2026Donghyun Lee, Jitesh Chavan, Duy Nguyen +5Diffusion TransformersQuantizer

  2. KroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers

    Jul 23, 2026Yann Bouquet, Alireza Khodamoradi, Kristof Denolf +1Post-Training QuantizationDiffusion Transformers

  3. DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers

    May 16, 2026Sayeh Sharify, Mahsa Salmani, Hesham MostafaTransformer ArchitecturesSingular Value Decomposition