cs.LGMay 11, 2026

A Theory of Time-Sensitive Language Generation: Sparse Hallucination Beats Mode Collapse

Authors: Atul GanjuTravis McVoyShaddin DughmiShang-Hua Teng

Organizations: University of Southern California

Abstract

We study language generation in the limit under a global preference ordering on strings, as introduced by Kleinberg and Wei. As is done in previous work, we aim for breadth, but impose an additional requirement of timeliness: higher-ranked strings should be generated earlier. A string is then only credited if it is generated before a deadline, where its deadline is defined by a function that maps a string's rank in the target language to the time by which it must be produced. This is in keeping with a central consideration in machine learning, where inductive bias favors simpler'' or more plausible'' outputs, all else being equal. We show that timely generation is impossible in a strong sense for eventually consistent generators -- the protagonists of most prior related work. Under what is perhaps the mildest natural relaxation of consistency, a hallucination rate that vanishes over time, we show that we can circumvent our impossibility result. In particular, we can achieve optimal density with respect to any superlinear deadline function. We also show this is tight by ruling out timely generation with linear deadlines and vanishing hallucination rate.

Explore similar work

CardsList
  1. Hallucination Rates in Language Generation

    Jul 25, 2026Debmalya Panigrahi, Fan Wei, Ian ZhangHallucination RateHallucinations

  2. Mistake-Bounded Language Generation

    May 11, 2026Jon Kleinberg, Charlotte Peale, Omer ReingoldX-Logsmask

  3. Space-Efficient Language Generation in the Limit

    Jun 24, 2026Nicolas Flammarion, Chirag Pabbaraju, Hristo Papazov +2Context-Free GrammarsData Generation