cs.CVOct 7, 2026

Do Generative Priors Align with Human Naturalness Perception?

Authors: Taiki Fukiage

Organizations: Communication Science Laboratories, NTT, Inc.

Abstract

Visual generative models are trained to capture the probability distributions of natural images, yet whether their native priors reflect the regularities governing human perception of image naturalness remains an open question. Here, we probe these priors through native prediction errors across 25 open image and video generators. Because raw single-image losses are dominated by scene content and visual complexity, we evaluate directional loss differences using content-preserving, paired relational interventions that selectively disrupt facial configurations or physical illumination consistency while limiting changes in low-level image statistics. Across both domains, these loss differences reproduce human-like selective sensitivities and tolerances, capturing the classic Thatcher effect on faces and shape-dependent responses to illumination inconsistencies. Notably, these loss differences reliably track continuous gradations of human naturalness judgments across individual stimulus pairs (peaking at r=.84r = .84 on faces and .64.64 on physical scenes) and retain unique human-aligned signals even after controlling for feature distances from frozen vision encoders and standard image quality metrics. We also find that while overall sensitivity to these violations broadly covaries with human alignment across models, the two systematically decouple along denoising schedules, with alignment peaking earlier than sensitivity, revealing that human-like naturalness judgments dissociate from generic violation detection. Together, these findings demonstrate that learning visual distributions yields generative loss landscapes that capture distinct aspects of human naturalness perception.

Figures & tables

Appendix figures & tables39 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Not Too Generative, Not Too Discriminative: The Human Alignment Sweet Spot

    May 22, 2026Jorge Chang Ortega, Bastien Le Lan, Thomas Serre +1Visual Representation LearningEnergy-Based Models

  2. Human face perception reflects inverse-generative and naturalistic discriminative objectives

    May 12, 2026Wenxuan Guo, Heiko H. Schütt, Kamila Maria Jozwik +3Face Recognition

  3. Threshold-Guided Optimization for Visual Generative Models

    May 6, 2026Jinbin Bai, Yu Lei, Qingyu Shi +6Diffusion Model AlignmentPreference Alignment