cs.AIOct 1, 2026

Gacha Decoding: Eliciting Diverse Generations Through Instruction Following

Authors: Scott Geng, Yufei Zhang, Joseph Lee, Jerry Li, Marjan Ghazvininejad, Pang Wei Koh

Organizations: University of Washington · Meta Superintelligence Labs

Abstract

We introduce Gacha Decoding, an inference-time method for eliciting diverse language model generations that scales with model capability. Across open-ended domains (in-the-wild chat, creative writing, planning for image generation, and protein design), Gacha Decoding significantly outperforms existing generation diversity approaches at equal quality (up to 2.4x Vendi over the next-best prior approach), reaching the same number of high-quality modes with over an order of magnitude fewer samples (11.0x) and discovering novel modes that no other approach surfaces. Our key insight is to treat diversity as an instruction-following problem: rather than relying on the LM's token entropy, we combine its instruction-following capability with randomness from an external RNG tool to scalably identify and realize distinct modes of the response space. This approach of "planning with dice" enables Gacha to invert the long-observed tension between diversity and model capability. As the underlying LM becomes a better instruction follower, diversity under Gacha Decoding consistently improves--even as its token entropy and diversity under prior approaches decline. Together, our results highlight that instruction following, rather than token entropy alone, can drive generation diversity.

Figures & tables

Appendix figures & tables6 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Where You Inject Diversity Matters: A Unified Framework for Diverse Generation

    Jun 9, 2026Cheng Zhang, Rui Xin, Chudi ZhongDiversityUnified Framework

  2. CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

    Aug 7, 2026Ananya Sahu, Mohit Bansal, Elias Stengel-EskinCreativityNarrative Generation

  3. Large Language Models Explore by Latent Distilling

    Apr 27, 2026Yuanhao Zeng, Ao Lu, Lufei Li +3DiversityProbe-Logit Distillation