vix.ing · top · new · best · stats · spec

Erik McDermott

  1. Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
    2020/02/07 by Qian Zhang, Zhang, Qian, Lu Han +11 · 15 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. A Density Ratio Approach to Language Model Fusion in End-To-End\n Automatic Speech Recognition
    2020/02/25 by Erik McDermott, Haşim Sak, McDermott, Erik +3 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  3. Optimizing Bilingual Neural Transducer with Synthetic Code-switching Text Generation
    2022/10/21 by Thien Huu Nguyen, Nathalie Tran, Nguyen, Thien +35 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  4. Optimizing Byte-level Representation for End-to-end ASR
    2024/06/14 by Roger Hsiao, Hsiao, Roger, Liuhui Deng +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning and Algorithms #electronic engineering #information engineering