Librispeech Speech Recognition
Librispeech is a widely used benchmark dataset for automatic speech recognition (ASR), driving advancements in speech processing. Current research focuses on improving ASR performance in challenging scenarios like multi-talker environments and noisy conditions, leveraging models such as Transformers, Conformers, and neural transducers, often incorporating techniques like self-supervised learning and knowledge distillation. These efforts aim to create more robust and accurate ASR systems, with implications for various applications including voice assistants, transcription services, and accessibility technologies. The development of larger datasets, such as Libriheavy, and the exploration of techniques like curriculum learning and multi-resolution processing further enhance the capabilities and efficiency of ASR models.
Papers
A Comparative Study on E-Branchformer vs Conformer in Speech Recognition, Translation, and Understanding Tasks
Yifan Peng, Kwangyoun Kim, Felix Wu, Brian Yan, Siddhant Arora, William Chen, Jiyang Tang, Suwon Shon, Prashant Sridhar, Shinji Watanabe
ZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMs
Xingchen Song, Di Wu, Binbin Zhang, Zhendong Peng, Bo Dang, Fuping Pan, Zhiyong Wu
Massively Multilingual ASR on 70 Languages: Tokenization, Architecture, and Generalization Capabilities
Andros Tjandra, Nayan Singhal, David Zhang, Ozlem Kalinli, Abdelrahman Mohamed, Duc Le, Michael L. Seltzer
Self-supervised learning with bi-label masked speech prediction for streaming multi-talker speech recognition
Zili Huang, Zhuo Chen, Naoyuki Kanda, Jian Wu, Yiming Wang, Jinyu Li, Takuya Yoshioka, Xiaofei Wang, Peidong Wang
Predicting Multi-Codebook Vector Quantization Indexes for Knowledge Distillation
Liyong Guo, Xiaoyu Yang, Quandong Wang, Yuxiang Kong, Zengwei Yao, Fan Cui, Fangjun Kuang, Wei Kang, Long Lin, Mingshuang Luo, Piotr Zelasko, Daniel Povey
Blank Collapse: Compressing CTC emission for the faster decoding
Minkyu Jung, Ohhyeok Kwon, Seunghyun Seo, Soonshin Seo