Emotional Speech Synthesis

Momentum

5 papers in the last four weeks, up 67% on the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 26

All topics
CardsList
  1. Steerspeech: Activation Steering For Emotion Control In Generated Speech

    Oct 7, 2026Afsara Benazir, Darius Pétermann, Felix Xiaozhu Lin +1Controllable Speech GenerationEmotional Speech Synthesis

  2. Tracing a Sparse Emotion-Control Circuit in LLM-Based Text-to-Speech

    Oct 4, 2026Hongfei Du, Jiacheng Shi, Yanfu Zhang +1Mechanistic InterpretabilityControllable Speech Generation

  3. EmoRES-TTS: Residual-Enhanced Vector Steering for Emotional Speech Generation

    Sep 29, 2026Kuan-Po Huang, Haohe Liu, Puyuan Peng +5Controllable Speech GenerationEmotional Speech Synthesis

  4. ReaFlow-TTS: Realization-Conditioned Flow Matching for High-Quality and Controllable Speech Synthesis

    Sep 24, 2026Junyi Zhao, Yihao Qin, Changsheng MaFlow MatchingTTS Synthesis

  5. Continuous-Time Acoustic Modelling with Neural Controlled Differential Equations

    Sep 12, 2026Mattias Cross, Minghui Zhao, Anton RagniNeural Controlled Differential EquationsTTS Synthesis

  6. Post-Training Zero-Shot TTS for Fine-Grained Emotion and Duration Control via Natural Language

    Sep 10, 2026Lianru Gao, Yujie Guo, Yong QinTTS SynthesisZero-Shot TTS

  7. Sequential Trajectories and Simultaneous Blending: Multi-Emotion Modeling for Instruction-Following TTS

    Aug 31, 2026Yan Zhou, Yun Hong, Yang FengControllable Speech GenerationEmotional Speech Synthesis

  8. Luna-TTS Family Technical Report

    Aug 12, 2026Feng Yin, Shuai Shi, Junjie Zheng +19Non-Autoregressive Text GenerationTTS Synthesis

  9. Let Me Look at You: Advanced Facial Expression Modeling for Conversational Speech Synthesis

    Jul 27, 2026Yifan Hu, Shuwei He, Rui Liu +1TTS SynthesisAudio-Driven Facial Animation

  10. AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

    Jul 17, 2026Zhenqi Jia, Yuan Zhao, Aruukhan +2TTS SynthesisEmotional Speech Synthesis

  11. A Geometric Perspective on Composable Emotion Steering in Text-to-Speech Models

    Jul 1, 2026Siyi Wang, James Bailey, Ting DangLanguage Model SteeringConditional Flow Matching

  12. LuxEmo: Expressive Text-to-Speech Corpus for Luxembourgish

    Jun 30, 2026Nina Hosseini-Kivanani, Sandipana DowerahSpeech ProcessingLow-Resource TTS Synthesis

  13. HPRO: Hierarchical Progressive Reward Optimization via Preference Extraction for Emotional Text-to-Speech

    Jun 26, 2026Sihang Nie, Xiaofen Xing, Rui Xing +5Reward ModelingSpeech Prosody

  14. Emo-LiPO: Listwise Preference Optimization for Fine-Grained Emotion Intensity Control in LLM-based Text-to-Speech

    Jun 11, 2026Yihang Lin, Li Zhou, Congwei Cao +4TTS SynthesisPreference Optimization

  15. EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis

    Jun 8, 2026Minghui Wu, Ganjun Liu, Zikun Fang +6Representation LearningTTS Synthesis

  16. Task-Vector Arithmetic for Emotional Expressivity Control in Language-Model-Based Text-to-Speech

    Jun 3, 2026Daniel Oliveira de Brito, Arnaldo Candido JuniorTTS SynthesisTask Arithmetic

  17. Sparse Autoencoders for Interpretable Emotion Control in Text-to-Speech

    May 31, 2026Hongfei Du, Jiacheng Shi, Sidi Lu +2TTS SynthesisSparse Autoencoders

  18. Sympatheia: Emotionally Adaptive Voice Assistant with Continuous Affect Conditioning

    May 30, 2026Sukru Samet Dindar, Riki Shimizu, Xilin Jiang +1Spoken Dialogue SystemsControllable Speech Generation

  19. PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

    May 26, 2026Bowen Li, Shaotong Guo, Zhen Wang +11TTS SynthesisZero-Shot TTS

  20. DUET: Unified Dual-Space Emotion Control for Diffusion and Flow-Matching Driven Text-to-Speech

    May 20, 2026Xu Zhang, Longbing Cao, Zhangkai WuTTS SynthesisAudio Diffusion Models

  21. The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

    Apr 29, 2026Yun-Shao Tsai, Yi-Cheng Lin, Huang-Cheng Chou +5TTS SynthesisSpeech Generation Evaluation

  22. ATRIE: Adaptive Tuning for Robust Inference and Emotion in Persona-Driven Speech Synthesis

    Apr 21, 2026Aoduo Li, Haoran Lv, Hongjian Xu +5TTS SynthesisControllable Speech Generation

  23. SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation

    Mar 23, 2026Lucas H. Ueda, João G. T. Lima, Pedro R. Corrêa +3Disentangled Representation LearningTTS Synthesis