vix.ing · top · new · best · stats · spec

Zhang, Xiao-Lei

  1. Speaker Recognition Based on Deep Learning: An Overview
    2020/12/02 by Zhongxin Bai, Xiao-Lei Zhang, Bai, Zhongxin +1 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Optimizing Quantum Federated Learning Based on Federated Quantum Natural Gradient Descent
    2023/02/27 by Jun Qi, Qi, Jun, Xiao-Lei Zhang +3 · 5 citations
    Computer Science · Engineering · Physics and Astronomy · #Advancements in Semiconductor Devices and Circuit Design #FOS: Computer and information sciences #FOS: Physical sciences #Machine Learning (cs.LG) #Quantum Computing Algorithms and Architecture #Quantum Physics (quant-ph) #Quantum and electron transport phenomena
  3. UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
    2025/02/06 by Lei Zhao, Linfeng Feng, Zhao, Lei +12 · 8 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Music Technology and Sound Studies
  4. Deep Ad-hoc Beamforming
    2018/11/03 by Xiao-Lei Zhang, Zhang, Xiao-Lei · 2 citations
    Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Millimeter-Wave Propagation and Modeling #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  5. Unsupervised model compression for multilayer bootstrap networks
    2015/03/22 by Zhang, Xiao-Lei · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
  6. Enhancing Intelligibility for Generative Target Speech Extraction via Joint Optimization with Target Speaker ASR
    2025/01/24 by Hao Ma, Ma, Hao, Rujin Chen +7 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Deep NMF Topic Modeling
    2021/02/24 by Jianyu Wang, Wang, JianYu, Xiao-Lei Zhang +1 · 2 citations
    Computer Science · Social Sciences · #Topic Modeling #Computational and Text Analysis Methods #Advanced Text Analysis Techniques
  8. Improving Pseudo Labels With Intra-Class Similarity for Unsupervised Domain Adaptation
    2022/07/25 by Wang, Jie, Zhang, Xiao-Lei · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  9. WeKws: A production first small-footprint end-to-end Keyword Spotting Toolkit
    2022/10/30 by Jie Wang, Wang, Jie, Menglong Xu +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #ICT in Developing Communities #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  10. Fast-U2++: Fast and Accurate End-to-End Speech Recognition in Joint CTC/Attention Frames
    2022/11/02 by Chengdong Liang, Xiao-Lei Zhang, Liang, Chengdong +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. DualSpec: Text-to-spatial-audio Generation via Dual-Spectrogram Guided Diffusion Model
    2025/02/26 by Lei Zhao, Zhao, Lei, Sizhou Chen +11 · 3 citations
    Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  12. Eliminating Quantization Errors in Classification-Based Sound Source Localization
    2023/11/21 by Feng, Linfeng, Zhang, Xiao-Lei, Li, Xuelong · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  13. FoleySpace: Vision-Aligned Binaural Spatial Audio Generation
    2025/08/18 by Lei Zhao, Zhao, Lei, Rujin Chen +7 · 2 citations
    Neuroscience · Computer Science · #Hearing Loss and Rehabilitation #Speech and Audio Processing #Generative Adversarial Networks and Image Synthesis