Zhenyao Zhu
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
2015/12/08 by Dario Amodei, Amodei, Dario, Rishita Anubhai +66 · 1 voice · 170 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL
- Deep Speaker: an End-to-End Neural Speaker Embedding System
2017/05/05 by Chao Li, Li, Chao, Xiaokong Ma +15 · 50 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
- Exploring Neural Transducers for End-to-End Speech Recognition
2017/07/24 by Eric Battenberg, Battenberg, Eric, Jitong Chen +19 · 7 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing
- Fully Supervised Speaker Diarization
2018/10/10 by Aonan Zhang, Quan Wang, Zhang, Aonan +7 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Gram-CTC: Automatic Unit Selection and Target Decomposition for Sequence Labelling
2017/03/01 by Hairong Liu, Zhenyao Zhu, Liu, Hairong +5 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis