vix.ing · top · new · best · stats · spec

Kwangyoun Kim

  1. E-Branchformer: Branchformer with Enhanced merging for speech recognition
    2022/09/30 by Kwangyoun Kim, Felix F. Wu, Kim, Kwangyoun +11 · 15 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  2. Improving ASR Contextual Biasing with Guided Attention
    2024/01/16 by Jiyang Tang, Tang, Jiyang, Kwangyoun Kim +9 · 7 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  3. SRU++: Pioneering Fast Recurrence with Attention for Speech Recognition
    2021/10/11 by Jing Pan, Pan, Jing, Tao Lei +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  4. DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
    2024/06/13 by Suwon Shon, Shon, Suwon, Kwangyoun Kim +9 · 1 citation
    Computer Science · #Speech and dialogue systems #Speech Recognition and Synthesis #Natural Language Processing Techniques
  5. Sample-Efficient Diffusion for Text-To-Speech Synthesis
    2024/09/01 by Justin Lovelace, Lovelace, Justin, Soham Ray +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems