vix.ing · top · new · best · stats · spec

Strimel, Grant P.

  1. Contextual Adapters for Personalized Speech Recognition in Neural Transducers
    2022/05/26 by Kanthashree Mysore Sathyendra, Thejaswi Muniyappa, Sathyendra, Kanthashree Mysore +13 · 4 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  2. Robust Acoustic and Semantic Contextual Biasing in Neural Transducers for Speech Recognition
    2023/05/09 by Fu, Xuandi, Sathyendra, Kanthashree Mysore, Gandhe, Ankur +4 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  3. SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning
    2025/04/12 by Pandey, Prabhat, Swaminathan, Rupak Vignesh, Girish, K V Vijay +4 · 7 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  4. Semantic Complexity in End-to-End Spoken Language Understanding
    2020/08/06 by McKenna, Joseph P., Choudhary, Samridhi, Saxon, Michael +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  5. CoDERT: Distilling Encoder Representations with Co-learning for\n Transducer-based Speech Recognition
    2021/06/14 by Rupak Vignesh Swaminathan, Brian King, Swaminathan, Rupak Vignesh +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  6. A neural prosody encoder for end-ro-end dialogue act classification
    2022/05/11 by Wei, Kai, Knox, Dillon, Radfar, Martin +6 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. Lookahead When It Matters: Adaptive Non-causal Transformers for Streaming Neural Transducers
    2023/05/07 by Grant P. Strimel, Yi Xie, Strimel, Grant P. +9 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  8. Dialog act guided contextual adapter for personalized speech recognition
    2023/03/31 by Chang, Feng-Ju, Muniyappa, Thejaswi, Sathyendra, Kanthashree Mysore +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering