vix.ing · top · new · best · stats · spec

Peng, Kainan

  1. Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence Learning
    2017/10/20 by Wei Ping, Ping, Wei, Kainan Peng +15 · 1 voice · 18 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.AI #cs.CL #cs.LG #cs.SD #eess.AS
  2. Neural Voice Cloning with a Few Samples
    2018/02/14 by Sercan Ö. Arık, Jitong Chen, Arik, Sercan O. +7 · 7 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. WaveFlow: A Compact Flow-based Model for Raw Audio
    2019/12/03 by Ping, Wei, Peng, Kainan, Zhao, Kexin +1 · 6 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. Deep Voice 2: Multi-Speaker Neural Text-to-Speech
    2017/05/24 by Sercan Ö. Arık, Arik, Sercan, Gregory Diamos +13 · 8 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  5. Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
    2025/02/11 by Zhang, Xueyao, Zhang, Xiaohui, Peng, Kainan +10 · 18 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  6. VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
    2024/04/10 by Philip Anastassiou, Anastassiou, Philip, Zhenyu Tang +15 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Zero-Shot Accent Conversion using Pseudo Siamese Disentanglement Network
    2022/12/12 by Dongya Jia, Qiao Tian, Jia, Dongya +13 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  8. ClariNet: Parallel Wave Generation in End-to-End Text-to-Speech
    2018/07/19 by Ping, Wei, Peng, Kainan, Chen, Jitong · 1 citation
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering