Dai, Lirong
- MoESys: A Distributed and Efficient Mixture-of-Experts Training and Inference System for Internet Services
2022/05/20 by Dianhai Yu, Yu, Dianhai, Liang Shen +11 · 8 citations
Computer Science · #Advanced Neural Network Applications #Recommender Systems and Techniques #Stochastic Gradient Optimization Techniques
- Deep-FSMN for Large Vocabulary Continuous Speech Recognition
2018/03/04 by Shiliang Zhang, Ming Lei, Zhang, Shiliang +5 · 6 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
2022/09/30 by Zhang, Ziqiang, Chen, Sanyuan, Zhou, Long +8 · 3 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
- LDM-SVC: Latent Diffusion Model Based Zero-Shot Any-to-Any Singing Voice Conversion with Singer Guidance
2024/06/08 by Shihao Chen, Yu Gu, Chen, Shihao +11 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Feedforward Sequential Memory Networks: A New Structure to Learn Long-term Dependency
2015/12/28 by Shiliang Zhang, Cong Liu, Zhang, Shiliang +9 · 1 citation
Computer Science · #FOS: Computer and information sciences #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
- Multi-Scale Attention with Dense Encoder for Handwritten Mathematical Expression Recognition
2018/01/05 by Zhang, Jianshu, Du, Jun, Dai, Lirong · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- SiFiSinger: A High-Fidelity End-to-End Singing Voice Synthesizer based on Source-filter Model
2024/10/16 by Cui, Jianwei, Gu, Yu, Weng, Chao +3 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
2022/03/31 by Ao, Junyi, Zhang, Ziqiang, Zhou, Long +7 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Multichannel AV-wav2vec2: A Framework for Learning Multichannel Multi-Modal Speech Representation
2024/01/07 by Zhu, Qiushi, Zhang, Jie, Gu, Yu +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Deep CLAS: Deep Contextual Listen, Attend and Spell
2024/09/26 by Mengzhi Wang, Wang, Mengzhi, Shifu Xiong +9 · 1 citation
Health Professions · Computer Science · #Interpreting and Communication in Healthcare #Speech and dialogue systems #Natural Language Processing Techniques
- CSSinger: End-to-End Chunkwise Streaming Singing Voice Synthesis System Based on Conditional Variational Autoencoder
2024/12/12 by Jianwei Cui, Yu Gu, Cui, Jianwei +9 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing