Speech Generation Evaluation

Latest papers 35

All topics
CardsList
  1. XSQ-AST: An Explainable Audio Spectrogram Transformer Framework for Localising Synthetic Speech Artifacts

    Sep 21, 2026Ben Heritage, Luca Resti, Mónica Villanueva Aylagas +3Explainable Artificial IntelligenceSpeech Generation Evaluation

  2. HaikuS2S: A Cascaded System For Responding In Verse

    Sep 20, 2026Devangi Sharma, Sophia Judicke, Glenda Tan +2TTS SynthesisSpeech Generation Evaluation

  3. Multi-Dimensional Prosody Judgment For Live Streaming Speech Synthesis

    Sep 17, 2026Zifan Guan, Longyu Lu, Junan Zhang +3Speech Generation EvaluationTTS Evaluation

  4. Quantifying the Generation Modality Gap in Speech-Text Language Models

    Sep 13, 2026Ju-Chieh Chou, Jiawei Zhou, Karen LivescuSpeech Generation EvaluationSpeech Language Models

  5. Voice or Stereotype? Disentangling Acoustic and Content-Based Gender in Speech-to-Speech Models

    Sep 10, 2026Xiaoqun Liu, Tanu Mitra, Harshit Rajgarhia +1Speech Generation EvaluationSpeech Processing

  6. MMAG: A Multi-Control Mixed Audio Generation Benchmark

    Aug 7, 2026Zihao Zheng, Xuenan Xu, Jiahao Mei +5Speech Generation EvaluationAudio Generation

  7. Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study

    Aug 3, 2026Ali Jafar, Amal Sarmad, Shifa Yousaf +1Speech Generation EvaluationTTS Evaluation

  8. Beyond Prompt Adherence: Auditing Attribute-Level Voice Control in Speech Generation

    Aug 1, 2026Xianhao Zhou, Jianghao WuSpeech Generation EvaluationTTS Evaluation

  9. A Production-Oriented Framework for Evaluation of SFX Generation

    Jul 10, 2026Mélodie Desbos, Yara Bahram, Eric Granger +1Speech Generation EvaluationAudio Editing

  10. SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models

    Jul 6, 2026Thomas Thebaud, Yuzhe Wang, Hao Zhang +5Audio-Language Model EvaluationLanguage Model Generation Evaluation

  11. Towards Digital Preservation of Efik: TTS for a Low-Resource African Language

    Jul 5, 2026Offiong Bassey Edet, Emmanuel Oyo-Ita, Archibong Okon Archibong +2TTS SynthesisSpeech Generation Evaluation

  12. Towards a Phonology-Informed Evaluation of Multilingual TTS

    Jul 2, 2026Sneha Ray Barman, Neeraj Kumar Sharma, Shakuntala MahantaTTS SynthesisSpeech Generation Evaluation

  13. Is Natural Always Appropriate? Investigating Naturalness and Appropriateness Across Different Domains for TTS Evaluation

    Jun 30, 2026Dominika Woszczyk, Andreas Triantafyllopoulos, Jura Miniota +2TTS SynthesisSpeech Generation Evaluation

  14. Reference-Based Prosody and Rhythm Evaluation for Spoken Dialogue Systems

    Jun 30, 2026Ashish Hallur, Thomas Thebaud, Georgi Tinchev +2Voice Agent EvaluationSpeech Generation Evaluation

  15. ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judge

    Jun 23, 2026Jisu Jeon, Seungyeon Jwa, Joosung Lee +6Audio-Language Model EvaluationPairwise Comparison

  16. An Evaluation Framework for Text-to-Speech Voice Reconstruction

    Jun 19, 2026Ariadna Sanchez, Christoph Minixhofer, Korin Richmond +3Speech Generation EvaluationTTS Evaluation

  17. Investigating Human-Model Discrepancies in Speech Quality Assessment via Acoustic and Prosodic Perturbations

    Jun 18, 2026Masato Takagi, Masaya Kawamura, Reo Shimizu +1Human Preference EvaluationSpeech Generation Evaluation

  18. AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

    Jun 2, 2026Haitao Li, Tian Tan, Yuguang Yang +2Audio-Language Model EvaluationLLM-as-a-Judge

  19. UniSRM: A Unified Speech Reward Model for Reasoning-Based Fine-grained Assessment

    May 22, 2026Yuanyuan Wang, Dongchao Yang, Yayue Deng +4Reward ModelingSpeech Generation Evaluation

  20. The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

    Apr 29, 2026Yun-Shao Tsai, Yi-Cheng Lin, Huang-Cheng Chou +5TTS SynthesisSpeech Generation Evaluation