vix.ing · top · new · best · stats · spec

Kang, Wei

  1. Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context
    2023/09/15 by Wei Kang, Xiaoyu Yang, Kang, Wei +12 · 33 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Zipformer: A faster and better encoder for automatic speech recognition
    2023/10/17 by Zengwei Yao, Yao, Zengwei, Liyong Guo +14 · 24 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
  3. PromptASR for contextualized ASR with controllable style
    2023/09/14 by Yang, Xiaoyu, Kang, Wei, Yao, Zengwei +5 · 8 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  4. Pruned RNN-T for fast, memory-efficient ASR training
    2022/06/23 by Fangjun Kuang, Liyong Guo, Kuang, Fangjun +10 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  5. Data-Driven Computational Methods for the Domain of Attraction and Zubov's Equation
    2021/12/29 by Kang, Wei, Sun, Kai, Xu, Liang · 5 citations
    #93-08 #Dynamical Systems (math.DS) #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Machine Learning (cs.LG) #Systems and Control (eess.SY) #electronic engineering #information engineering
  6. ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
    2025/06/16 by Zhu, Han, Kang, Wei, Yao, Zengwei +6 · 11 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. Capacity of a Class of Diamond Channels
    2008/08/07 by Kang, Wei, Ulukus, Sennur · 1 citation
    #FOS: Computer and information sciences #H.1.1 #Information Theory (cs.IT)
  8. The Gaussian Multiple Access Diamond Channel
    2011/04/17 by Kang, Wei, Liu, Nan, Chong, Weiwei · 1 citation
    #FOS: Computer and information sciences #Information Theory (cs.IT)
  9. k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
    2024/11/26 by Yifan Yang, Yang, Yifan, Jianheng Zhuo +20 · 4 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  10. Predicting Multi-Codebook Vector Quantization Indexes for Knowledge Distillation
    2022/10/31 by Liyong Guo, Xiaoyu Yang, Guo, Liyong +20 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  11. CR-CTC: Consistency regularization on CTC for improved speech recognition
    2024/10/07 by Yao, Zengwei, Kang, Wei, Yang, Xiaoyu +7 · 4 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  12. Capacity of Hierarchical Secure Coded Gradient Aggregation with Straggling Communication Links
    2024/12/16 by Qian Nan Lu, Lu, Qinyi, Jiale Cheng +5 · 4 citations
    Computer Science · Engineering · #Advanced Wireless Communication Technologies #Cooperative Communication and Network Coding #FOS: Computer and information sciences #Information Theory (cs.IT) #Wireless Communication Security Techniques
  13. The Capacity of Symmetric Private Information Retrieval under Arbitrary Collusion and Eavesdropping Patterns
    2020/10/16 by Jiale Cheng, Nan Liu, Cheng, Jiale +3 · 1 citation
    Computer Science · #Complexity and Algorithms in Graphs #Cryptography and Data Security #FOS: Computer and information sciences #Information Theory (cs.IT) #Optimization and Search Problems
  14. LibriheavyMix: A 20,000-Hour Dataset for Single-Channel Reverberant Multi-Talker Speech Separation, ASR and Speaker Diarization
    2024/09/01 by Zengrui Jin, Yifan Yang, Jin, Zengrui +22 · 2 citations
    Computer Science · Health Professions · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Infant Health and Development #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  15. Fast and parallel decoding for transducer
    2022/10/31 by Wei Kang, Kang, Wei, Liyong Guo +14 · 1 citation
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #DNA and Biological Computing #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Network Packet Processing and Optimization #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  16. Blank-regularized CTC for Frame Skipping in Neural Transducer
    2023/05/19 by Yifan Yang, Yang, Yifan, Xiaoyu Yang +14 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Neural Networks and Applications #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  17. Delay-penalized CTC implemented based on Finite State Transducer
    2023/05/19 by Yao, Zengwei, Kang, Wei, Kuang, Fangjun +5 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  18. ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching
    2025/07/12 by Zhu, Han, Kang, Wei, Guo, Liyong +10 · 4 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering