cs.CLSep 30, 2026

Mixture of Decoders for Diverse Dialog Response Generation

Authors: Wenchao Du

Organizations: Microsoft Corporation Redmond, WA 98052

Abstract

Mixture modeling is a long established machine learning technique for learning large sets of multi-modal data. While it is known that sequence-to-sequence models for dialog response generation suffer from the problem of low diversity, we hypothesize that it is because sequence-to-sequence models tend to learn a degenerate uni-modal distribution of responses. We then propose to incorporate a mixture of decoders into sequence-to-sequence models and try to make each decoder learn specialized topics in order to improve the diversity of generated responses. Our model is developed under the framework of conditional variational autoencoder (CVAE). We evaluate our approach on an open domain chat corpus and show improvement over strong baselines in quantitative measures and human evaluation.

Figures & tables

Explore similar work

CardsList
  1. Gacha Decoding: Eliciting Diverse Generations Through Instruction Following

    Oct 1, 2026Scott Geng, Yufei Zhang, Joseph Lee +3Large Language Model GenerationDiversity

  2. Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses

    May 27, 2026Kerui Peng, Feifei Li, Xingyu Fan +1Large Language Model GenerationResponses

  3. Where You Inject Diversity Matters: A Unified Framework for Diverse Generation

    Jun 9, 2026Cheng Zhang, Rui Xin, Chudi ZhongDiversityUnified Framework