cs.CVSep 30, 2026

D-Scope: Decomposing and Steering Diffusion Transformers with Sparse Autoencoders

Authors: Xinyue Xu, Jiahao Zhang, Lijie Hu, Peter Hase, Hao Wang

Organizations: Pivotal Research · Mohamed bin Zayed University of Artificial Intelligence · Schmidt Sciences · Stanford University · University of Illinois at Urbana-Champaign

Abstract

Sparse autoencoders (SAEs) reveal visual structure in diffusion transformers (DiTs), but interpreting a feature does not establish whether it can be used to control generation. We introduce D-Scope (Diffusion Scope), a framework that connects feature interpretation to generation control through shared visual evidence. D-Scope aggregates SigLIP2 embeddings of highly activating image patches into visual centroids. Matching target text descriptions against these visual centroids in the shared image-text embedding space then enables retrieval of individual features without per-feature text annotations. The underlying patches provide evidence for inspecting each selection, while spatially masked interventions test the corresponding decoder direction at varying strengths under fixed generation conditions. We characterize 150 SAEs across two model families and five layers, and introduce a benchmark of 100 target concepts with ten contexts each spanning under-specified and explicit-conflict conditions. Our empirical results show that high reconstruction fidelity can coexist with low dictionary utilization and limited visual-evidence coverage. Under per-case best-of-sweep strength selection, contrastive retrieval yields larger mean regional SigLIP2 gains than direct retrieval across the tested steering configurations, without consistently improving outside-region preservation. D-Scope provides an inspectable framework for evaluating sparse DiT features through their visual evidence and the effects of their decoder directions on generation. The demo is available at https://jiahaozhang-public.github.io/d-scope/.

Figures & tables

Appendix figures & tables26 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models

    Jun 30, 2026Enrico Cassano, Riccardo Renzulli, Rayyan Ahmed +2Diffusion ModelsConcept Erasure

  2. Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers

    May 28, 2026Zihao Xue, Yan Wang, Zhen Bi +7Diffusion TransformersTransformer Architectures

  3. Awakening Diffusion Transformers: Eliciting Stronger Generation and Understanding via Massive Activation Modulation

    Jul 3, 2026Chaofan Gan, Zicheng Zhao, Yuanpeng Tu +6Diffusion TransformersRepresentational Capacity