cs.LGOct 7, 2026

Self-Consuming Generative Models with Co-Evolving Human Preferences

Authors: Xiukun Wei, Tian Xie, Ding Zhu, Xueru Zhang

Organizations: The Ohio State University

Abstract

Generative models are increasingly trained in self-consuming iterative loops, where users curate preferred samples from model-generated candidates and the curated samples are used to train future generations of the model. Prior work has largely assumed fixed user preferences, but in practice exposure to model outputs gradually reshapes what users perceive as desirable, creating a feedback loop in which model distributions and user preferences co-evolve. We take a first step toward understanding the long-term behavior of such coupled dynamics. We show that when training relies entirely on user-curated synthetic data, iterative curation amplifies initial biases and drives the system toward one of multiple singleton equilibria in which the instance holding an initial advantage eventually dominates. In contrast, injecting reference data into training at a sufficiently large rate fundamentally changes the dynamics and yields a unique globally attracting equilibrium. Building on this insight, we study how reference-data injection can be used to control long-term outcomes, and propose an efficient algorithm that jointly selects a reference distribution and its mixing weight to steer the coupled system toward equilibria that preserve desired attributes while minimizing data collection costs.

Figures & tables

Appendix figures & tables11 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences

    May 8, 2026Ali Falahati, Mohammad Mohammadi Amiri, Kate Larson +1Synthetic DataGenerative Models

  2. When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

    May 28, 2026Yang Zhang, Xiukun Wei, Xueru ZhangPreference AlignmentReinforcement Learning From Human Feedback

  3. Stability and Diversity of Networked Self-Consuming Generative Ecosystems

    Oct 7, 2026Xiukun Wei, Yang Zhang, Xueru Zhang