David Rybach
- Streaming End-to-end Speech Recognition For Mobile Devices
2018/11/15 by Yanzhang He, He, Yanzhang, Tara N. Sainath +37 · 9 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
- Lingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling
2019/02/21 by Jonathan Shen, Shen, Jonathan, Patrick Nguyen +179 · 1 voice · 2 citations
Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.LG #stat.ML
- A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency
2020/03/28 by Tara N. Sainath, Yanzhang He, Sainath, Tara N. +55 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing
- E2E Segmenter: Joint Segmenting and Decoding for Long-Form ASR
2022/04/22 by W. Ronny Huang, Shuo-Yiin Chang, Huang, W. Ronny +13 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Personalized Speech recognition on mobile devices
2016/03/10 by Ian McGraw, Rohit Prabhavalkar, McGraw, Ian +22 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.SD
- Mobile Keyboard Input Decoding with Finite-State Transducers
2017/04/13 by Tom Ouyang, David Rybach, Ouyang, Tom +5 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
- Less Is More: Improved RNN-T Decoding Using Limited Label Context and\n Path Merging
2020/12/12 by Rohit Prabhavalkar, Yanzhang He, Prabhavalkar, Rohit +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering