vix.ing · top · new · best · stats · spec

Guo, Jinxi

  1. Prompting Large Language Models with Speech Recognition Abilities
    2023/07/21 by Yassir Fathullah, Fathullah, Yassir, Chunyang Wu +21 · 23 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. A spelling correction model for end-to-end speech recognition
    2019/02/19 by Jinxi Guo, Guo, Jinxi, Tara N. Sainath +3 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. Singing voice conversion with non-parallel data
    2019/03/11 by Xin Chen, Wei Chu, Chen, Xin +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Variable frame rate-based data augmentation to handle speaking-style variability for automatic speaker verification
    2020/08/08 by Afshan, Amber, Guo, Jinxi, Park, Soo Jin +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Signal Processing (eess.SP) #electronic engineering #information engineering
  5. Improving Fast-slow Encoder based Transducer with Streaming Deliberation
    2022/12/15 by Li, Ke, Mahadeokar, Jay, Guo, Jinxi +5 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  6. Effective internal language model training and fusion for factorized transducer model
    2024/04/02 by Guo, Jinxi, Moritz, Niko, Ma, Yingyi +6 · 1 citation
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #electronic engineering #information engineering
  7. Transducer-Llama: Integrating LLMs into Streamable Transducer-based Speech Recognition
    2024/12/21 by Deng, Keqi, Guo, Jinxi, Ma, Yingyi +4 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering