cs.CLOct 1, 2026

No Model Required: Text Entropy Rate Filtering Mitigates Iterative Fine-Tuning Collapse

Authors: Lewis Mitchell

Organizations: Adelaide Data Science Centre School of Mathematical Sciences Adelaide University Adelaide SA 5005, Australia

Abstract

Iterative fine-tuning on synthetic data causes \emph{model collapse}: output diversity narrows as rare patterns are progressively lost, a signature most visible as phrase-level repetition. Existing mitigations either require model log-probabilities, an external oracle, or continued access to real human data. Here we develop a new approach grounded in mathematical information theory: the non-parametric Kontoyiannis entropy rate estimator hkh_k, computed entirely from raw text via match-length statistics, with no model of any kind. We show that this is in fact a \emph{superior} training-data filter on text-diversity metrics in a fully-synthetic, single-lineage fine-tuning setting. In a six-generation QLoRA collapse experiment on Llama-3.1-8B, logprob-based filtering (the most established model-access-requiring baseline) provides no significant text-diversity benefit on any metric (p>0.23p > 0.23), whereas hkh_k-filtering yields +42%+42\% unique trigrams, +30%+30\% vocabulary, and −19%-19\% repetition (all p<0.001p < 0.001). We validate hkh_k as a cross-domain entropy proxy (β=0.924β= 0.924, R2=0.746R^2 = 0.746) and collapse detector (ρ=+0.454ρ= +0.454, p<0.0001p < 0.0001) across 4domains, 2temperatures, 2~generator--scorer model pairs, and 1{,}520 generated documents. Our results demonstrate that information theoretic approaches to collapse mitigation are efficient, and suggest new approaches for maintaining multi-agent diversity.

Figures & tables

Appendix figures & tables4 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Learning from Synthetic Data without Model Collapse in Iterative Instruction Tuning

    Jul 19, 2026Xiaonan Luo, Yue Huang, Kehan Guo +4Instruction TuningSynthetic Training Data

  2. Learning by Surprise: Adaptive Mitigation of Model Collapse in Large Language Models

    Oct 16, 2024Daniele Gambetta, Gizem Gezici, Fosca Giannotti +3CollapseGenerative Artificial Intelligence

  3. Fine-Tuning Improves Information Conveyance in Language Models

    May 29, 2026Yuwei Cheng, Weiyi Tian, Haifeng XuLarge Language Model Fine-TuningModel Fine-Tuning