vix.ing · top · new · best · stats · spec

Zhenyao Zhu

  1. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
    2015/12/08 by Dario Amodei, Amodei, Dario, Rishita Anubhai +66 · 1 voice · 170 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL
  2. Deep Speaker: an End-to-End Neural Speaker Embedding System
    2017/05/05 by Chao Li, Li, Chao, Xiaokong Ma +15 · 50 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  3. Exploring Neural Transducers for End-to-End Speech Recognition
    2017/07/24 by Eric Battenberg, Battenberg, Eric, Jitong Chen +19 · 7 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing
  4. Fully Supervised Speaker Diarization
    2018/10/10 by Aonan Zhang, Quan Wang, Zhang, Aonan +7 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Gram-CTC: Automatic Unit Selection and Target Decomposition for Sequence Labelling
    2017/03/01 by Hairong Liu, Zhenyao Zhu, Liu, Hairong +5 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis