Ling, Zhenhua
- The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods
2018/04/12 by Jaime Lorenzo-Trueba, Junichi Yamagishi, Lorenzo-Trueba, Jaime +11 · 17 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- Voice Conversion Challenge 2020: Intra-lingual semi-parallel and cross-lingual voice conversion
2020/08/28 by Yi Zhao, Wen-Chin Huang, Zhao, Yi +13 · 12 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Topic Modeling #Voice and Speech Disorders
- Predictions of Subjective Ratings and Spoofing Assessments of Voice Conversion Challenge 2020 Submissions
2020/09/08 by Rohan Kumar Das, Tomi Kinnunen, Das, Rohan Kumar +13 · 3 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- SyncSpeech: Low-Latency and Efficient Dual-Stream Text-to-Speech based on Temporal Masked Transformer
2025/02/16 by Sheng, Zhengyan, Du, Zhihao, Zhang, Shiliang +3 · 6 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Sound (cs.SD)
- DiffStyleTTS: Diffusion-based Hierarchical Prosody Modeling for Text-to-Speech with Diverse and Controllable Styles
2024/12/04 by Jiaxuan Liu, Liu, Jiaxuan, Zhaoci Liu +9 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
- Using multiple reference audios and style embedding constraints for speech synthesis
2021/10/09 by Gong, Cheng, Wang, Longbiao, Ling, Zhenhua +2 · 1 citation
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Refining Self-Supervised Learnt Speech Representation using Brain Activations
2024/06/12 by Hengyu Li, Kangdi Mei, Li, Hengyu +11 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Intelligent Tutoring Systems and Adaptive Learning #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering