vix.ing · top · new · best · stats · spec

Felix F. Wu

  1. E-Branchformer: Branchformer with Enhanced merging for speech recognition
    2022/09/30 by Kwangyoun Kim, Kim, Kwangyoun, Felix F. Wu +11 · 16 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  2. SLUE: New Benchmark Tasks for Spoken Language Understanding Evaluation\n on Natural Speech
    2021/11/19 by Suwon Shon, Shon, Suwon, Ankita Pasad +11 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  3. Improving ASR Contextual Biasing with Guided Attention
    2024/01/16 by Jiyang Tang, Kwangyoun Kim, Tang, Jiyang +9 · 7 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  4. Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
    2021/09/14 by Felix F. Wu, Kwangyoun Kim, Wu, Felix +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  5. Sample-Efficient Diffusion for Text-To-Speech Synthesis
    2024/09/01 by Justin Lovelace, Soham Ray, Lovelace, Justin +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems