Jinglin Liu
- AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
2023/04/25 by Rongjie Huang, Huang, Rongjie, Mingze Li +23 · 1 voice · 52 citations
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Topic Modeling
- Make-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models
2023/01/30 by Rongjie Huang, Huang, Rongjie, Jiawei Huang +17 · 43 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
2021/05/06 by Jinglin Liu, Liu, Jinglin, Chengxi Li +7 · 25 citations
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Music Technology and Sound Studies
- GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis
2023/01/31 by Zhenhui Ye, Ziyue Karen Jiang, Ye, Zhenhui +9 · 22 citations
Computer Science · #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
- Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation
2023/05/29 by Jiawei Huang, Huang, Jiawei, Yi Ren +17 · 21 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus
2021/12/20 by Rongjie Huang, Feiyang Chen, Huang, Rongjie +9 · 14 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- ProDiff: Progressive Fast Diffusion Model For High-Quality Text-to-Speech
2022/07/13 by Rongjie Huang, Huang, Rongjie, Zhou Zhao +9 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis
2023/07/14 by Ziyue Karen Jiang, Jinglin Liu, Jiang, Ziyue +21 · 9 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation
2023/05/01 by Zhenhui Ye, Jinzheng He, Ye, Zhenhui +17 · 8 citations
Computer Science · #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
- RMSSinger: Realistic-Music-Score based Singing Voice Synthesis
2023/05/18 by Jinzheng He, He, Jinzheng, Jinglin Liu +11 · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
2023/06/06 by Ziyue Karen Jiang, Yi Ren, Jiang, Ziyue +21 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- PortaSpeech: Portable and High-Quality Generative Text-to-Speech
2021/09/30 by Yi Ren, Ren, Yi, Jinglin Liu +3 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- CLAPSpeech: Learning Prosody from Text Context with Contrastive Language-Audio Pre-training
2023/05/18 by Zhenhui Ye, Ye, Zhenhui, Rongjie Huang +13 · 4 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- A Study of Non-autoregressive Model for Sequence Generation
2020/04/22 by Yi Ren, Ren, Yi, Jinglin Liu +9 · 2 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
- VarietySound: Timbre-Controllable Video to Sound Generation via Unsupervised Information Disentanglement
2022/11/19 by Chenye Cui, Yi Ren, Cui, Chenye +7 · 2 citations
Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech and Audio Processing
- DopplerBAS: Binaural Audio Synthesis Addressing Doppler Effect
2022/12/14 by Jinglin Liu, Zhenhui Ye, Liu, Jinglin +11 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
- Ada-TTA: Towards Adaptive High-Quality Text-to-Talking Avatar Synthesis
2023/06/06 by Zhenhui Ye, Ye, Zhenhui, Ziyue Karen Jiang +13 · 2 citations
Computer Science · #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
- DenoiSpeech: Denoising Text to Speech with Frame-Level Noise Modeling
2020/12/17 by Chen Zhang, Zhang, Chen, Yi Ren +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- GenerSpeech: Towards Style Transfer for Generalizable Out-Of-Domain Text-to-Speech
2022/05/15 by Rongjie Huang, Yi Ren, Huang, Rongjie +7 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing