Voice Cloning

Momentum

5 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 24

All topics
CardsList
  1. RAWD-TTS: Ratio-Free Reward Alignment for Discrete-Diffusion Voice Cloning

    Sep 29, 2026Maxim Maslov, Kirill Borodin, Vasilii Kudryavtsev +2Audio Diffusion ModelsZero-Shot TTS

  2. From Reliable Text to Real Voices: Trust-Aware Progressive Adaptation for Low-Resource TTS

    Sep 22, 2026Jiayi Lu, Yizhong Geng, Jinghan Yang +4Synthetic-to-Real Domain AdaptationLow-Resource TTS Synthesis

  3. Taming Long-form Text-to-Speech

    Sep 15, 2026Rongxiang Wang, Berkin Durmus, Aysegul Orhon +2TTS SynthesisTTS Evaluation

  4. Cross-Lingual F5-TTS 2: A Simplified Framework for Language-Agnostic Voice Cloning

    Sep 14, 2026Qingyu Liu, Rixi Xu, Yushen Chen +11TTS SynthesisZero-Shot TTS

  5. CookVoice: Unified Framework for Style Controllable Multi-Modal Human Voice Generation

    Aug 12, 2026Haowei Lou, Hye-Young Paik, Dai Jia +2Singing Voice SynthesisTTS Synthesis

  6. Extracting Voice Styles from Frozen TTS Models via Gradient-Based Inverse Optimization

    Jul 28, 2026Gyeongmin KimTTS SynthesisVoice Cloning

  7. Synthetic Speech, Real Signal: Paralinguistic Preservation and Cross-Lingual Augmentation via Voice Cloning

    Jul 24, 2026Roseline Polle, Owen Parsons, George Fairs +5Synthetic Data AugmentationSpeech Processing

  8. ZONOS2 Technical Report

    Jun 23, 2026Gabriel Clark, Sofian Mejjoute, Mohamed Osman +2TTS SynthesisStreaming TTS Synthesis

  9. Dynamic Prosody Prediction in LLM-based TTS for Improving Speaker Similarity

    Jun 13, 2026Zhenwei Mou, Liping Chen, Yajun Hu +3TTS SynthesisSpeech Prosody

  10. Vocal Identity Under Siege by AI Voice Cloning Technologies

    Jun 11, 2026Jyh-An Lee, Xuan SunAudio GenerationVoice Cloning

  11. Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation

    Jun 5, 2026Ziyu Zhang, Chunyu Qiang, Xiaopeng Wang +10Controllable Music GenerationMusic Generation

  12. VoxCPM2 Technical Report

    Jun 5, 2026Yixuan Zhou, Guoyang Zeng, Xin Liu +15Speech Foundation ModelsControllable Speech Generation

  13. Voice "Cloning" is Style Transfer

    May 15, 2026Kaitlyn Zhou, Federico Bianchi, Martijn Bartelds +3Controllable Speech GenerationTrust in AI

  14. X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning

    May 7, 2026Rixi Xu, Qingyu Liu, Haitao Li +10TTS SynthesisZero-Shot TTS

  15. JaiTTS: A Thai Voice Cloning Model

    Apr 30, 2026Jullajak Karnjanaekarin, Pontakorn Trakuekul, Narongkorn Panitsrisit +5TTS SynthesisAudio Generation

  16. One Voice, Many Tongues: Cross-Lingual Voice Cloning for Scientific Speech

    Apr 28, 2026Amanuel Gizachew Abebe, Yasmin MoslemTTS SynthesisText-to-Speech

  17. V.O.I.C.E (Voice, Ownership, Identity, Control, Expression): Risk Taxonomy of Synthetic Voice Generation From Empirical Data

    Apr 25, 2026Tanusree Sharma, Anish Krishnagiri, Lili Dudas +2TTS SynthesisAudio Generation

  18. Acoustic and perceptual differences between standard and accented speech and their voice clones

    Apr 2, 2026Tianle Yang, Chengzhe Sun, Phil Rose +1TTS SynthesisAudio Generation

  19. Speech Generation Speaker Poisoning: Capability Erasure in Zero-Shot Text-to-Speech

    Mar 8, 2026Thanathai Lertpetchpun, Thanapat Trachu, Sai Praneeth Karimireddy +1TTS SynthesisZero-Shot TTS

  20. RVCBench: Benchmarking the Robustness of Voice Cloning Across Modern Audio Generation Models

    Jan 31, 2026Ruinan Jin, Xinting Liao, Hanlin Yu +2Adversarial RobustnessAudio Generation