vix.ing · top · new · best · stats · spec

Zhong Meng

  1. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
    2023/03/02 by Yu Zhang, Wei Han, Zhang, Yu +55 · 1 voice · 46 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.SD #eess.AS #electronic engineering #information engineering
  2. Continuous speech separation: dataset and analysis
    2020/01/30 by Zhuo Chen, Chen, Zhuo, Takuya Yoshioka +15 · 18 citations
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  3. Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers
    2020/06/19 by Naoyuki Kanda, Kanda, Naoyuki, Yashesh Gaur +11 · 10 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. SLM: Bridge the thin gap between speech and text foundation models
    2023/09/30 by Mingqiu Wang, Wei Han, Wang, Mingqiu +33 · 13 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  5. Streaming Speaker-Attributed ASR with Token-Level Speaker Embeddings
    2022/03/30 by Naoyuki Kanda, Kanda, Naoyuki, Wu, Jian +16 · 7 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Transcribe-to-Diarize: Neural Speaker Diarization for Unlimited Number of Speakers using End-to-End Speaker-Attributed ASR
    2021/10/07 by Naoyuki Kanda, Xiong Xiao, Kanda, Naoyuki +11 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Factorized Neural Transducer for Efficient Language Model Adaptation
    2021/09/27 by Xie Chen, Zhong Meng, Chen, Xie +5 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Music and Audio Processing
  8. Contextual Biasing with the Knuth-Morris-Pratt Matching Algorithm
    2023/09/29 by Weiran Wang, Zelin Wu, Wang, Weiran +23 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  9. Continuous Speech Separation with Ad Hoc Microphone Arrays
    2021/03/03 by Dongmei Wang, Takuya Yoshioka, Wang, Dongmei +9 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Continuous Speech Separation with Recurrent Selective Attention Network
    2021/10/28 by Yixuan Zhang, Zhang, Yixuan, Zhuo Chen +10 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Investigation of End-To-End Speaker-Attributed ASR for Continuous Multi-Talker Recordings
    2020/08/11 by Naoyuki Kanda, Xuankai Chang, Kanda, Naoyuki +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  12. Augmenting conformers with structured state-space sequence models for online speech recognition
    2023/09/15 by Haozhe Shan, Albert Gu, Shan, Haozhe +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  13. High-Accuracy and Low-Latency Speech Recognition with Two-Head Contextual Layer Trajectory LSTM Model
    2020/03/17 by Jinyu Li, Li, Jinyu, Rui Zhao +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Efficiently Train ASR Models that Memorize Less and Perform Better with Per-core Clipping
    2024/06/04 by Lun Wang, Om Thakkar, Wang, Lun +9 · 1 citation
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Fault Detection and Control Systems #Machine Learning and ELM #Sound (cs.SD) #electronic engineering #information engineering