F5-Tts

Momentum

5 papers in the last four weeks, against 2 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 20

All topics
CardsList
  1. RVQ Position Aware Speculative Decoding for On Device Text to Speech

    Sep 29, 2026Berkin Durmus, Eduardo Pacheco, Zach Nagengast +1Autoregressive Text-To-SpeechAutoregressive Decoding

  2. ReaFlow-TTS: Realization-Conditioned Flow Matching for High-Quality and Controllable Speech Synthesis

    Sep 24, 2026Junyi Zhao, Yihao Qin, Changsheng MaFlow-Matching Text-To-SpeechSpeech Synthesis

  3. TTS-Guard: Black-Box Ownership Verification of Text-to-Speech Models via Adaptive Adversarial Speaker-Pair Fingerprints

    Sep 20, 2026Xubin Yue, Zhenhua Xu, Zhebo Wang +5Cross-Lingual Voice CloningF5-Tts

  4. Cross-Lingual F5-TTS 2: A Simplified Framework for Language-Agnostic Voice Cloning

    Sep 14, 2026Qingyu Liu, Rixi Xu, Yushen Chen +11Cross-Lingual Voice CloningF5-Tts

  5. KABURI-TTS: Phoneme-Keyed Activity-conditioned Bi-channel Utterance Rendering for Interaction

    Sep 7, 2026Ryuichiro Higashinaka, Shinnosuke Takamichi, Tetsuji OgawaFull-Duplex Speech ModelsF5-Tts

  6. Luna-TTS Family Technical Report

    Aug 12, 2026Feng Yin, Shuai Shi, Junjie Zheng +19Autoregressive Text-To-SpeechF5-Tts

  7. FreyaTTS Technical Report

    Jul 10, 2026Ahmet Erdem Pamuk, Ömer Yentür, Ahmet Tunga Bayrak +2F5-TtsGrapheme-To-Phoneme

  8. E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation

    Jun 25, 2026Wen Ye, Peiyan Li, Tingyu Yuan +7Robotic ManipulationEmbodied

  9. ISCSLP 2026 CoT-TTS Challenge: Chain-of-Thought Reasoning for Context-Aware Text-to-Speech

    Jun 20, 2026Wei Xue, Junlan Feng, Shilei Zhang +9F5-TtsChain-of-Thought Reasoning

  10. Streaming T5-based Text-to-Speech Synthesis with Limited Lookahead

    Jun 20, 2026Muyang Du, Jason Roche, Junjie LaiText-To-Speech SynthesisF5-Tts

  11. FineCombo-TTS: Collaborative and Precise Controllable Speech Synthesis Using Text Descriptions and Reference Speech

    Jun 17, 2026Shuoyi Zhou, Yixuan Zhou, Peiji Yang +4Speech SynthesisF5-Tts

  12. EmoInstruct-TTS: Dual-Path Instruction-Guided Emotional Speech Synthesis

    Jun 8, 2026Minghui Wu, Ganjun Liu, Zikun Fang +6Speech SynthesisEmotion

  13. Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech

    Jun 5, 2026Adarsh Arigala, Arjun Gangwar, S Umesh +1Flow-Matching Text-To-SpeechSpeech Synthesis

  14. X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning

    May 7, 2026Rixi Xu, Qingyu Liu, Haitao Li +10Cross-Lingual Voice CloningMultilingual Automatic Speech Recognition

  15. MAGIC-TTS: Fine-Grained Controllable Speech Synthesis with Explicit Local Duration and Pause Control

    Apr 23, 2026Jialong Mai, Xiaofen Xing, Xiangmin XuF5-TtsSpeech Synthesis

  16. CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment

    Feb 23, 2026Hanwen Liu, Saierdaer Yusuyin, Hao Huang +1F5-TtsSpeech-To-Text Alignment