Chen, Zhengyang
- Wespeaker: A Research and Production oriented Speaker Embedding Learning Toolkit
2022/10/31 by Hongji Wang, Wang, Hongji, Chengdong Liang +13 · 32 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- Large-scale Self-Supervised Speech Representation Learning for Automatic Speaker Verification
2021/10/12 by Zhengyang Chen, Chen, Zhengyang, Sanyuan Chen +13 · 23 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training
2021/10/12 by Sanyuan Chen, Chen, Sanyuan, Yu Wu +18 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
2024/07/21 by Shuai Wang, Wang, Shuai, Zhengyang Chen +7 · 9 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Attention-based Encoder-Decoder End-to-End Neural Diarization with Embedding Enhancer
2023/09/13 by Zhengyang Chen, Chen, Zhengyang, Bing Han +5 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Attention-based Encoder-Decoder Network for End-to-End Neural Speaker Diarization with Target Speaker Attractor
2023/05/18 by Zhengyang Chen, Bing Han, Chen, Zhengyang +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Self-Supervised Learning with Cluster-Aware-DINO for High-Performance Robust Speaker Verification
2023/04/12 by Bing Han, Zhengyang Chen, Han, Bing +3 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Self-Supervised Speaker Verification Using Dynamic Loss-Gate and Label Correction
2022/08/03 by Bing Han, Zhengyang Chen, Han, Bing +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling
2025/07/22 by Guo, Yi, Wang, Wei, Yuan, Zhihang +8 · 5 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
2024/06/13 by Chen, Zhengyang, Liu, Xuechen, Cooper, Erica +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Flow-TSVAD: Target-Speaker Voice Activity Detection via Latent Flow Matching
2024/09/07 by Zhengyang Chen, Bing Han, Chen, Zhengyang +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Disentangling the Prosody and Semantic Information with Pre-trained Model for In-Context Learning based Zero-Shot Voice Conversion
2024/09/08 by Zhengyang Chen, Chen, Zhengyang, Shuai Wang +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Advanced Zero-Shot Text-to-Speech for Background Removal and Preservation with Controllable Masked Speech Prediction
2025/02/11 by Leying Zhang, Wangyou Zhang, Zhang, Leying +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering