cs.CVSep 29, 2026

The Domain Is a Residue: Adapting Self-Supervised Features, Not Generators

Authors: Thomas Deixelberger, Markus Steinberger

Organizations: Huawei, Austria · Graz University of Technology, Austria

Abstract

Clearing fog, rain or snow from footage, or turning renders into photographs, must remove the source domain and keep the scene. Unpaired translators carry it through because their generator sees the source appearance (pixels, a near-invertible latent or a control map) and keeps it. A DINO feature map fixes what is in the scene and carries weather, lighting and rendering style as a residue of 13 to 14% of the feature norm. We propose the Representation Feature Adapter (RFA), a 2.9M-parameter network that moves this residue. We train only the adapter and its discriminators; the encoder and a feature-conditioned decoder, trained once for all conditions, stay frozen. Against CycleGAN-Turbo it is ahead on both metrics on fog and on KID on night, and level within noise on snow, rain and haze. On sim-to-real it leads REGEN and HyPER-GAN on both metrics. Only the RFA removes the rain while keeping the scene. The removal costs scene structure: CycleGAN-Turbo keeps more on every condition but fog. On VAE latents the identical adapter collapses to the identity, and decoders from other groups that never saw it render its output. The RFA has about 160 times fewer trainable parameters than CycleGAN-Turbo and under a fifth of its per-condition training time.

Figures & tables

Appendix figures & tables18 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion

    Jun 18, 2026Aryan Das, Koushik Biswas, Moloud Abdar +1Text-Conditioned Diffusion ModelImage Generation

  2. Generative Manifold Distillation: Aligning Restoration Trajectories with Natural Image Prior

    Dec 11, 2025Yuyang Hu, Mojtaba Sahraee-Ardakan, Arpit Bansal +4Image Restoration