vix.ing · top · new · best · stats · spec

Chen, Zhengyang

  1. Wespeaker: A Research and Production oriented Speaker Embedding Learning Toolkit
    2022/10/31 by Hongji Wang, Wang, Hongji, Chengdong Liang +13 · 32 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
  2. Large-scale Self-Supervised Speech Representation Learning for Automatic Speaker Verification
    2021/10/12 by Zhengyang Chen, Chen, Zhengyang, Sanyuan Chen +13 · 23 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
    2021/10/12 by Sanyuan Chen, Chen, Sanyuan, Yu Wu +18 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
    2024/07/21 by Shuai Wang, Wang, Shuai, Zhengyang Chen +7 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Attention-based Encoder-Decoder End-to-End Neural Diarization with Embedding Enhancer
    2023/09/13 by Zhengyang Chen, Chen, Zhengyang, Bing Han +5 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Attention-based Encoder-Decoder Network for End-to-End Neural Speaker Diarization with Target Speaker Attractor
    2023/05/18 by Zhengyang Chen, Bing Han, Chen, Zhengyang +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Self-Supervised Learning with Cluster-Aware-DINO for High-Performance Robust Speaker Verification
    2023/04/12 by Bing Han, Zhengyang Chen, Han, Bing +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Self-Supervised Speaker Verification Using Dynamic Loss-Gate and Label Correction
    2022/08/03 by Bing Han, Zhengyang Chen, Han, Bing +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling
    2025/07/22 by Guo, Yi, Wang, Wei, Yuan, Zhihang +8 · 5 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  10. Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
    2024/06/13 by Chen, Zhengyang, Liu, Xuechen, Cooper, Erica +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. Flow-TSVAD: Target-Speaker Voice Activity Detection via Latent Flow Matching
    2024/09/07 by Zhengyang Chen, Bing Han, Chen, Zhengyang +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  12. Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
    2024/09/08 by Zhengyang Chen, Chen, Zhengyang, Shuai Wang +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  13. Advanced Zero-Shot Text-to-Speech for Background Removal and Preservation with Controllable Masked Speech Prediction
    2025/02/11 by Leying Zhang, Wangyou Zhang, Zhang, Leying +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering