Zweig, Geoffrey
- Language Models for Image Captioning: The Quirks and What Works
2015/05/07 by Jacob Devlin, Devlin, Jacob, Hao Cheng +13 · 1 voice · 1 citation
#cs.CL #cs.AI #cs.CV #cs.LG
- From Captions to Visual Concepts and Back
2014/11/18 by Hao Fang, Fang, Hao, Saurabh Gupta +21 · 11 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications
- Contextual RNN-T For Open Domain ASR
2020/06/04 by Jain, Mahaveer, Keren, Gil, Mahadeokar, Jay +3 · 10 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Attention with Intention for a Neural Network Conversation Model
2015/10/29 by Kaisheng Yao, Yao, Kaisheng, Geoffrey Zweig +3 · 12 citations
Computer Science · #Topic Modeling #Speech and dialogue systems #Natural Language Processing Techniques
- Sequence-to-Sequence Neural Net Models for Grapheme-to-Phoneme Conversion
2015/05/31 by Kaisheng Yao, Yao, Kaisheng, Geoffrey Zweig +1 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- Multilingual Graphemic Hybrid ASR with Massive Data Augmentation
2019/09/14 by Chunxi Liu, Qiaochu Zhang, Liu, Chunxi +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Contextualizing ASR Lattice Rescoring with Hybrid Pointer Network Language Model
2020/05/15 by Liu, Da-Rong, Liu, Chunxi, Zhang, Frank +3 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Kaizen: Continuously improving teacher using Exponential Moving Average\n for semi-supervised speech recognition
2021/06/14 by Vimal Manohar, Manohar, Vimal, Tatiana Likhomanenko +13 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- An Attentional Neural Conversation Model with Improved Specificity
2016/06/03 by Kaisheng Yao, Baolin Peng, Yao, Kaisheng +5 · 2 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
- End-to-end LSTM-based dialog control optimized with supervised and reinforcement learning
2016/06/03 by Williams, Jason D., Zweig, Geoffrey · 1 citation
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Hybrid Code Networks: practical and efficient end-to-end dialog control with supervised and reinforcement learning
2017/02/10 by Williams, Jason D., Asadi, Kavosh, Zweig, Geoffrey · 1 citation
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Deja-vu: Double Feature Presentation and Iterated Loss in Deep\n Transformer Networks
2019/10/22 by Andros Tjandra, Tjandra, Andros, Chunxi Liu +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition
2019/10/02 by Le, Duc, Zhang, Xiaohui, Zheng, Weiyi +3 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- Large scale weakly and semi-supervised learning for low-resource video ASR
2020/05/16 by Singh, Kritika, Manohar, Vimal, Xiao, Alex +7 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering