vix.ing · top · new · best · stats · spec

David Rybach

  1. Streaming End-to-end Speech Recognition For Mobile Devices
    2018/11/15 by Yanzhang He, He, Yanzhang, Tara N. Sainath +37 · 9 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  2. Lingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling
    2019/02/21 by Jonathan Shen, Shen, Jonathan, Patrick Nguyen +179 · 1 voice · 2 citations
    Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.LG #stat.ML
  3. A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
    2020/03/28 by Tara N. Sainath, Yanzhang He, Sainath, Tara N. +55 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing
  4. E2E Segmenter: Joint Segmenting and Decoding for Long-Form ASR
    2022/04/22 by W. Ronny Huang, Shuo-Yiin Chang, Huang, W. Ronny +13 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Personalized Speech recognition on mobile devices
    2016/03/10 by Ian McGraw, Rohit Prabhavalkar, McGraw, Ian +22 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.SD
  6. Mobile Keyboard Input Decoding with Finite-State Transducers
    2017/04/13 by Tom Ouyang, David Rybach, Ouyang, Tom +5 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  7. Less Is More: Improved RNN-T Decoding Using Limited Label Context and\n Path Merging
    2020/12/12 by Rohit Prabhavalkar, Yanzhang He, Prabhavalkar, Rohit +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering