Erik McDermott
- Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
2020/02/07 by Qian Zhang, Zhang, Qian, Lu Han +11 · 15 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- A Density Ratio Approach to Language Model Fusion in End-To-End\n Automatic Speech Recognition
2020/02/25 by Erik McDermott, Haşim Sak, McDermott, Erik +3 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Optimizing Bilingual Neural Transducer with Synthetic Code-switching Text Generation
2022/10/21 by Thien Huu Nguyen, Nathalie Tran, Nguyen, Thien +35 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
- Optimizing Byte-level Representation for End-to-end ASR
2024/06/14 by Roger Hsiao, Hsiao, Roger, Liuhui Deng +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning and Algorithms #electronic engineering #information engineering