vix.ing · top · new · best · stats · spec

Chan, Julian

  1. Contextualized Streaming End-to-End Speech Recognition with Trie-Based Deep Biasing and Shallow Fusion
    2021/04/05 by Duc Le, Mahaveer Jain, Le, Duc +21 · 15 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition
    2020/10/21 by Yangyang Shi, Yongqiang Wang, Shi, Yangyang +13 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Deep Shallow Fusion for RNN-T Personalization
    2020/11/16 by Le, Duc, Keren, Gil, Chan, Julian +3 · 6 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  4. Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
    2025/02/11 by Xueyao Zhang, Xiaohui Zhang, Zhang, Xueyao +22 · 20 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Transformer in action: a comparative study of transformer-based acoustic models for large scale speech recognition applications
    2020/10/27 by Wang, Yongqiang, Shi, Yangyang, Zhang, Frank +4 · 2 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Sound (cs.SD)
  6. Just ASK: Building an Architecture for Extensible Self-Service Spoken Language Understanding
    2017/11/01 by Kumar, Anjishnu, Gupta, Arpit, Chan, Julian +9 · 1 citation
    #68T50 #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Neural and Evolutionary Computing (cs.NE) #Software Engineering (cs.SE)
  7. Dynamic Encoder Transducer: A Flexible Solution For Trading Off Accuracy For Latency
    2021/04/05 by Yangyang Shi, Shi, Yangyang, Varun Nagaraja +21 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing