vix.ing · top · new · best · stats · spec

Zweig, Geoffrey

  1. Language Models for Image Captioning: The Quirks and What Works
    2015/05/07 by Jacob Devlin, Devlin, Jacob, Hao Cheng +13 · 1 voice · 1 citation
    #cs.CL #cs.AI #cs.CV #cs.LG
  2. From Captions to Visual Concepts and Back
    2014/11/18 by Hao Fang, Fang, Hao, Saurabh Gupta +21 · 11 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  3. Contextual RNN-T For Open Domain ASR
    2020/06/04 by Jain, Mahaveer, Keren, Gil, Mahadeokar, Jay +3 · 10 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. Attention with Intention for a Neural Network Conversation Model
    2015/10/29 by Kaisheng Yao, Yao, Kaisheng, Geoffrey Zweig +3 · 12 citations
    Computer Science · #Topic Modeling #Speech and dialogue systems #Natural Language Processing Techniques
  5. Sequence-to-Sequence Neural Net Models for Grapheme-to-Phoneme Conversion
    2015/05/31 by Kaisheng Yao, Yao, Kaisheng, Geoffrey Zweig +1 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  6. Multilingual Graphemic Hybrid ASR with Massive Data Augmentation
    2019/09/14 by Chunxi Liu, Qiaochu Zhang, Liu, Chunxi +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Contextualizing ASR Lattice Rescoring with Hybrid Pointer Network Language Model
    2020/05/15 by Liu, Da-Rong, Liu, Chunxi, Zhang, Frank +3 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Kaizen: Continuously improving teacher using Exponential Moving Average\n for semi-supervised speech recognition
    2021/06/14 by Vimal Manohar, Manohar, Vimal, Tatiana Likhomanenko +13 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. An Attentional Neural Conversation Model with Improved Specificity
    2016/06/03 by Kaisheng Yao, Baolin Peng, Yao, Kaisheng +5 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
  10. End-to-end LSTM-based dialog control optimized with supervised and reinforcement learning
    2016/06/03 by Williams, Jason D., Zweig, Geoffrey · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  11. Hybrid Code Networks: practical and efficient end-to-end dialog control with supervised and reinforcement learning
    2017/02/10 by Williams, Jason D., Asadi, Kavosh, Zweig, Geoffrey · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  12. Deja-vu: Double Feature Presentation and Iterated Loss in Deep\n Transformer Networks
    2019/10/22 by Andros Tjandra, Tjandra, Andros, Chunxi Liu +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  13. From Senones to Chenones: Tied Context-Dependent Graphemes for Hybrid Speech Recognition
    2019/10/02 by Le, Duc, Zhang, Xiaohui, Zheng, Weiyi +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  14. Large scale weakly and semi-supervised learning for low-resource video ASR
    2020/05/16 by Singh, Kritika, Manohar, Vimal, Xiao, Alex +7 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering