vix.ing · top · new · best · stats · spec

Ke Hu

  1. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
    2023/03/02 by Yu Zhang, Zhang, Yu, Wei Han +51 · 27 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Deliberation Model Based Two-Pass End-to-End Speech Recognition
    2020/03/17 by Ke Hu, Hu, Ke, Tara N. Sainath +5 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Organellar dynamics during the cell cycle of Toxoplasma gondii
    2008/04/15 by Manami Nishi, Ke Hu, John M. Murray +1 · 26 citations
    Immunology and Microbiology · Medicine · #Toxoplasma gondii Research Studies #Parasitic Infections and Diagnostics #Herpesvirus Infections and Treatments
  4. Chain-of-Thought Prompting for Speech Translation
    2024/09/17 by Ke Hu, Hu, Ke, Zhehuai Chen +13 · 6 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
  5. SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
    2025/05/21 by Ke Hu, Ehsan Hosseini-Asl, Hu, Ke +17 · 11 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
  6. Learning Word-Level Confidence For Subword End-to-End ASR
    2021/03/11 by David Qiu, Qiu, David, Qiujia Li +21 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  7. Deliberation of Streaming RNN-Transducer by Non-autoregressive Decoding
    2021/12/01 by Weiran Wang, Wang, Weiran, Ke Hu +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Enhancing Visual Continual Learning with Language-Guided Supervision
    2024/03/24 by Bolin Ni, Ni, Bolin, Zhao, Hongbo +10 · 2 citations
    Social Sciences · #Reflective Practices in Education
  9. Intense high-harmonic optical vortices generated from a micro-plasma-waveguide irradiated by a circularly polarized laser pulse
    2022/02/17 by Ke Hu, Longqing Yi, Hu, Ke +1 · 1 citation
    Physics and Astronomy · #Advanced Fiber Laser Technologies #FOS: Physical sciences #Laser-Matter Interactions and Applications #Optics (physics.optics) #Orbital Angular Momentum in Optics #Plasma Physics (physics.plasm-ph)
  10. NeKo: Cross-Modality Post-Recognition Error Correction with Tasks-Guided Mixture-of-Experts Language Model
    2024/11/08 by Yen‐Ting Lin, Zhehuai Chen, Lin, Yen-Ting +23 · 2 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  11. EMMeTT: Efficient Multimodal Machine Translation Training
    2024/09/20 by Piotr Żelasko, Żelasko, Piotr, Zhehuai Chen +17 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  12. VoiceTextBlender: Augmenting Large Language Models with Speech Capabilities via Single-Stage Joint Speech-Text Supervised Fine-Tuning
    2024/10/23 by Yifan Peng, Peng, Yifan, Krishna C. Puvvada +17 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering