Autoregressive Text-To-Speech

Latest papers 24

All topics
CardsList
  1. RVQ Position Aware Speculative Decoding for On Device Text to Speech

    Sep 29, 2026Berkin Durmus, Eduardo Pacheco, Zach Nagengast +1Autoregressive Text-To-SpeechAutoregressive Decoding

  2. Controlling Speaking Rate in Autoregressive TTS via Activation Steering

    Sep 27, 2026Francesco Verdini, Antonis Asonitis, Aref Farhadipour +4Autoregressive Text-To-SpeechSeed-Tts-Eval Benchmark

  3. Taming Long-form Text-to-Speech

    Sep 15, 2026Rongxiang Wang, Berkin Durmus, Aysegul Orhon +2Autoregressive Text-To-SpeechVoice Conversion

  4. Continuous-Time Acoustic Modelling with Neural Controlled Differential Equations

    Sep 12, 2026Mattias Cross, Minghui Zhao, Anton RagniAutoregressive Text-To-SpeechNeural Audio

  5. Luna-TTS Family Technical Report

    Aug 12, 2026Feng Yin, Shuai Shi, Junjie Zheng +19Autoregressive Text-To-SpeechF5-Tts

  6. CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents

    Aug 9, 2026Yuqian Zhang, Yao Shi, Kexin Huang +6Autoregressive Text-To-SpeechCross-Lingual Voice Cloning

  7. Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens

    Jul 31, 2026Yi Luo, Rongzhi Gu, Jixun YaoAutoregressive Text-To-Speech

  8. StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis

    Jul 22, 2026Kaicheng Luo, Xuefei Gong, Yutao Sun +6Autoregressive Text-To-SpeechSpeech Synthesis

  9. Fréchet Distance Loss on Speech Representations for Text-to-Speech Synthesis

    Jul 7, 2026Ho-Lam Chung, Kuan-Po Huang, Bo-Ru Lu +1Flow-Matching Text-To-SpeechAutoregressive Text-To-Speech

  10. DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

    Jul 5, 2026Junwon Moon, Seungbeom Kim, Yejin Lee +4Autoregressive Text-To-SpeechDiffusion Language Models

  11. Probing Low Frame Rate Degradation in Neural Audio Codecs

    Jun 15, 2026Alex Gichamba, Moise BusogiNeural Audio CodecsFrame Rate

  12. TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech

    Jun 8, 2026Yejin Lee, Junwon Moon, Hyoeun Kim +3Autoregressive Text-To-SpeechAutoregressive Decoding

  13. Read What You Hear: Reference-Free Hypotheses Evaluation with Acoustic Discrepancy

    Jun 3, 2026Zhihan Li, Hankun Wang, Yiwei Guo +3Autoregressive Text-To-SpeechAcoustic

  14. PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

    May 26, 2026Bowen Li, Shaotong Guo, Zhen Wang +11Autoregressive Text-To-SpeechSpeech Synthesis

  15. JaiTTS: A Thai Voice Cloning Model

    Apr 30, 2026Jullajak Karnjanaekarin, Pontakorn Trakuekul, Narongkorn Panitsrisit +5Cross-Lingual Voice CloningAutoregressive Text-To-Speech

  16. StarTSE: Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model

    Apr 21, 2026Shuhai Peng, Hui Lu, Jinjiang Liu +8Target Speaker ExtractionAutoregressive Text-To-Speech

  17. AST: Adaptive, Seamless, and Training-Free Precise Speech Editing

    Apr 17, 2026Sihan Lv, Yechen Jin, Zhen Li +5Autoregressive Text-To-SpeechSyntax

  18. WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

    Mar 17, 2026Hanna Lee, Tan Dat Nguyen, Jaehoon Kang +1Autoregressive Text-To-SpeechKnowledge Distillation

  19. Decoding Order Matters in Autoregressive Speech Synthesis

    Jan 13, 2026Minghui Zhao, Anton RagniAutoregressive Text-To-SpeechAutoregressive Decoding