TTS Synthesis

TTS: Text-to-Speech

Momentum

41 papers in the last four weeks, up 242% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 180

All topics
CardsList
  1. Beyond Speech Captions: Speech-Rewarded Style Planning for Conversational Text-to-Speech

    Oct 8, 2026Shiao Zhu, Lianbo Liu, Sizhen Lyu +3TTS SynthesisControllable Speech Generation

  2. Edit Who Speaks, Control How They Speak: Global Timbre Editing and Local Instruction Control for TTS

    Oct 8, 2026Junchuan Zhao, Chenglin Xu, Wei Zeng +3TTS SynthesisControllable Speech Generation

  3. Steerspeech: Activation Steering For Emotion Control In Generated Speech

    Oct 7, 2026Afsara Benazir, Darius Pétermann, Felix Xiaozhu Lin +1Controllable Speech GenerationEmotional Speech Synthesis

  4. Training-Free Instruction TTS Gender Bias Calibration Using Model-Adaptive Steering

    Oct 7, 2026Kuan-Yu Chen, Yi-Cheng Lin, Jeng-Lin Li +1Text-to-SpeechTTS Synthesis

  5. Loud and Clear: Dynamic Activation Steering for Improving Speech Intelligibility in Noisy Environments

    Oct 6, 2026Seymanur Akti, Alexander WaibelTTS SynthesisControllable Speech Generation

  6. Pronunciation-Oriented Reinforcement Learning for Japanese Text-to-Speech with Kana-Domain ASR Rewards

    Oct 6, 2026Shiao Zhu, Lianbo Liu, Kai Washizaki +2TTS Synthesis

  7. Can Prosodic Style Be Inferred from Text Alone? Evidence from Unsupervised Acoustic Clusters

    Oct 4, 2026Abdul Rehman, Jian-Jun Zhang, Xiaosong YangTTS SynthesisSpeech Processing

  8. Balalaika-Longform: A Russian Speech Corpus for Continuous Long-Form Text-to-Speech

    Sep 30, 2026Nikita Vasiliev, Kirill Borodin, Vasilii Kudryavtsev +2TTS SynthesisTTS Evaluation

  9. SCIC: Scope- and Codebook-Aware Instruction Conditioning for Speaker-Adapted Expressive TTS

    Sep 30, 2026Longyu Lu, Zongwei Du, Mengtao Xing +4TTS SynthesisStreaming TTS Synthesis

  10. RVQ Position Aware Speculative Decoding for On Device Text to Speech

    Sep 29, 2026Berkin Durmus, Eduardo Pacheco, Zach Nagengast +1TTS SynthesisStreaming TTS Synthesis

  11. Repetition, Not Length: Isolating the Counting Failure in Neural Text-to-Speech

    Sep 29, 2026Kirill Borodin, Vasilii Kudryavtsev, Maxim Maslov +1TTS SynthesisTTS Evaluation

  12. Distill Locally, Schedule Globally: Flow Maps for Few-Step Text-to-Speech

    Sep 28, 2026Yentl Collin, Evan Dufraisse, Amr Mohamed +3Flow MatchingTTS Synthesis

  13. InstCharVoice: Grounding Natural-Language Instructions for Character-Level Control in Text-to-Speech

    Sep 28, 2026Sihang Nie, Xueru Li, Xiaofen Xing +4TTS SynthesisLanguage Model-Based Control

  14. SEmoEdit: Probing and Harnessing the Editability of Pre-trained Speech Flows

    Sep 28, 2026Tianxin Xie, Pengfei Zhang, Kai Jiang +2TTS SynthesisSpeech Editing

  15. Harmonizing Spectral Evolution in Conditional Flow Matching for TTS

    Sep 28, 2026Isha Pandey, Varad Deshpande, Abhijat Bharadwaj +1TTS SynthesisConditional Flow Matching

  16. Controlling Speaking Rate in Autoregressive TTS via Activation Steering

    Sep 27, 2026Francesco Verdini, Antonis Asonitis, Aref Farhadipour +4TTS SynthesisControllable Speech Generation

  17. DEFINE: Exemplar-Guided Accent Control for Zero-Shot TTS

    Sep 26, 2026Ambuj Mehrish, Abhinaba Roy, Alex Ivanov +2TTS SynthesisZero-Shot TTS

  18. EditVoice: Variable-Length Non-Autoregressive Zero-Shot TTS and Speech Editing with Edit Flows

    Sep 24, 2026Hongyao Deng, Wenhao Guan, Xuetao Lin +4TTS SynthesisSpeech Editing

  19. ReaFlow-TTS: Realization-Conditioned Flow Matching for High-Quality and Controllable Speech Synthesis

    Sep 24, 2026Junyi Zhao, Yihao Qin, Changsheng MaFlow MatchingTTS Synthesis

  20. EmphTTS: an emphasis-control TTS with reinforcement learning

    Sep 23, 2026Zirui Li, Rech Silas, Lauri Juvela +2TTS SynthesisControllable Speech Generation

  21. Not Quite My Tempo: Voice Activity-aware Speech Synthesis for Lip-Synchronous Dubbing

    Sep 22, 2026Alejandro Pérez-González-de-Martos, Florian Lux, Angelina Elizarova +3TTS SynthesisAudio-Visual Synchronization

  22. CycleSpeech: Reciprocal Alignment for Instruction-Controlled Speech Synthesis and Paralinguistic Understanding

    Sep 21, 2026Huan Liao, Haonan Han, Xingwen Han +3TTS SynthesisControllable Speech Generation