Peng, Kainan
- Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence Learning
2017/10/20 by Wei Ping, Ping, Wei, Kainan Peng +15 · 1 voice · 18 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.AI #cs.CL #cs.LG #cs.SD #eess.AS
- Neural Voice Cloning with a Few Samples
2018/02/14 by Sercan Ö. Arık, Jitong Chen, Arik, Sercan O. +7 · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- WaveFlow: A Compact Flow-based Model for Raw Audio
2019/12/03 by Ping, Wei, Peng, Kainan, Zhao, Kexin +1 · 6 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Deep Voice 2: Multi-Speaker Neural Text-to-Speech
2017/05/24 by Sercan Ö. Arık, Arik, Sercan, Gregory Diamos +13 · 8 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
2025/02/11 by Zhang, Xueyao, Zhang, Xiaohui, Peng, Kainan +10 · 18 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
2024/04/10 by Philip Anastassiou, Anastassiou, Philip, Zhenyu Tang +15 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Zero-Shot Accent Conversion using Pseudo Siamese Disentanglement Network
2022/12/12 by Dongya Jia, Qiao Tian, Jia, Dongya +13 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- ClariNet: Parallel Wave Generation in End-to-End Text-to-Speech
2018/07/19 by Ping, Wei, Peng, Kainan, Chen, Jitong · 1 citation
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering