Steve Renals
- The MGB-2 Challenge: Arabic Multi-Dialect Broadcast Media Recognition
2016/09/19 by Ahmed Ali, Ali, Ahmed, Peter Bell +11 · 4 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
- TaL: a synchronised multi-speaker corpus of ultrasound tongue imaging,\n audio, and lip videos
2020/11/19 by Manuel Sam Ribeiro, Ribeiro, Manuel Sam, Jennifer Sanger +11 · 4 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- On the Usefulness of Self-Attention for Automatic Speech Recognition with Transformers
2020/11/08 by Shucong Zhang, Erfan Loweimi, Zhang, Shucong +5 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Knowledge Distillation for Small-footprint Highway Networks
2016/08/02 by Liang Lu, Lu, Liang, Michelle Guo +3 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Segmental Recurrent Neural Networks for End-to-end Speech Recognition
2016/03/01 by Liang Lu, Lingpeng Kong, Lu, Liang +7 · 7 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.NE
- Lattice-Based Unsupervised Test-Time Adaptation of Neural Network Acoustic Models
2019/06/27 by Ondřej Klejch, Joachim Fainberg, Klejch, Ondrej +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Automatic Dialect Detection in Arabic Broadcast Speech
2015/09/23 by Ahmed Ali, Najim Dehak, Ali, Ahmed +13 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.CL
- Leveraging speaker attribute information using multi task learning for\n speaker verification and diarization
2020/10/27 by Chau Luu, Peter Bell, Luu, Chau +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Accessing the spoken word
2005/05/12 by Jerry Goldman, Steve Renals, Steven Bird +9 · 1 citation
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing