Felix F. Wu
- E-Branchformer: Branchformer with Enhanced merging for speech recognition
2022/09/30 by Kwangyoun Kim, Kim, Kwangyoun, Felix F. Wu +11 · 16 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- SLUE: New Benchmark Tasks for Spoken Language Understanding Evaluation\n on Natural Speech
2021/11/19 by Suwon Shon, Shon, Suwon, Ankita Pasad +11 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
- Improving ASR Contextual Biasing with Guided Attention
2024/01/16 by Jiyang Tang, Kwangyoun Kim, Tang, Jiyang +9 · 7 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
2021/09/14 by Felix F. Wu, Kwangyoun Kim, Wu, Felix +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Sample-Efficient Diffusion for Text-To-Speech Synthesis
2024/09/01 by Justin Lovelace, Soham Ray, Lovelace, Justin +7 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems