vix.ing · top · new · best · stats · spec

Hu, Ke

  1. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
    2023/03/02 by Yu Zhang, Zhang, Yu, Wei Han +51 · 26 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Diffusion-based Reinforcement Learning via Q-weighted Variational Policy Optimization
    2024/05/25 by Ding, Shutong, Hu, Ke, Zhang, Zhenhao +5 · 17 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  3. Deliberation Model Based Two-Pass End-to-End Speech Recognition
    2020/03/17 by Hu, Ke, Sainath, Tara N., Pang, Ruoming +1 · 6 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  4. Attaining the Unattainable? Reassessing Claims of Human Parity in Neural Machine Translation
    2018/08/30 by Toral, Antonio, Castilho, Sheila, Hu, Ke +1 · 3 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  5. Chain-of-Thought Prompting for Speech Translation
    2024/09/17 by Ke Hu, Zhehuai Chen, Hu, Ke +13 · 6 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
  6. SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
    2025/05/21 by Ke Hu, Hu, Ke, Ehsan Hosseini-Asl +17 · 11 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
  7. Learning Word-Level Confidence For Subword End-to-End ASR
    2021/03/11 by Qiu, David, Li, Qiujia, He, Yanzhang +9 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #electronic engineering #information engineering
  8. Deliberation of Streaming RNN-Transducer by Non-autoregressive Decoding
    2021/12/01 by Weiran Wang, Ke Hu, Wang, Weiran +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Multilingual and Fully Non-Autoregressive ASR with Large Language Model Fusion: A Comprehensive Study
    2024/01/23 by Huang, W. Ronny, Allauzen, Cyril, Chen, Tongzhou +7 · 3 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  10. Scaling Up Deliberation for Multilingual ASR
    2022/10/11 by Hu, Ke, Li, Bo, Sainath, Tara N. · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. Training and Inference Efficiency of Encoder-Decoder Speech Models
    2025/03/07 by Żelasko, Piotr, Dhawan, Kunal, Galvez, Daniel +7 · 6 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  12. Phoneme-Based Contextualization for Cross-Lingual Speech Recognition in End-to-End Models
    2019/06/21 by Hu, Ke, Bruguier, Antoine, Sainath, Tara N. +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  13. Enhancing Visual Continual Learning with Language-Guided Supervision
    2024/03/24 by Bolin Ni, Ni, Bolin, Zhao, Hongbo +10 · 2 citations
    Social Sciences · #Reflective Practices in Education
  14. Transformer Based Deliberation for Two-Pass Speech Recognition
    2021/01/27 by Hu, Ke, Pang, Ruoming, Sainath, Tara N. +1 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  15. Mixture-of-Expert Conformer for Streaming Multilingual ASR
    2023/05/25 by Hu, Ke, Li, Bo, Sainath, Tara N. +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  16. Intense high-harmonic optical vortices generated from a micro-plasma-waveguide irradiated by a circularly polarized laser pulse
    2022/02/17 by Ke Hu, Longqing Yi, Hu, Ke +1 · 1 citation
    Physics and Astronomy · #Advanced Fiber Laser Technologies #FOS: Physical sciences #Laser-Matter Interactions and Applications #Optics (physics.optics) #Orbital Angular Momentum in Optics #Plasma Physics (physics.plasm-ph)
  17. NeKo: Cross-Modality Post-Recognition Error Correction with Tasks-Guided Mixture-of-Experts Language Model
    2024/11/08 by Yen‐Ting Lin, Lin, Yen-Ting, Zhehuai Chen +23 · 2 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  18. EMMeTT: Efficient Multimodal Machine Translation Training
    2024/09/20 by Piotr Żelasko, Żelasko, Piotr, Zhehuai Chen +17 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  19. VoiceTextBlender: Augmenting Large Language Models with Speech Capabilities via Single-Stage Joint Speech-Text Supervised Fine-Tuning
    2024/10/23 by Yifan Peng, Peng, Yifan, Krishna C. Puvvada +17 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  20. GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
    2025/05/24 by Ding, Shutong, Hu, Ke, Zhong, Shan +5 · 2 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  21. Exploring the Boundary of Diffusion-based Methods for Solving Constrained Optimization
    2025/02/14 by Ding, Shutong, Zhou, Yimiao, Hu, Ke +4 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  22. KD-MSLRT: Lightweight Sign Language Recognition Model Based on Mediapipe and 3D to 1D Knowledge Distillation
    2025/01/04 by Li, Yulong, Ren, Bolin, Hu, Ke +4 · 1 citation
    #Computers and Society (cs.CY) #FOS: Computer and information sciences