vix.ing · top · new · best · stats · spec

Dai, Lirong

  1. MoESys: A Distributed and Efficient Mixture-of-Experts Training and Inference System for Internet Services
    2022/05/20 by Dianhai Yu, Yu, Dianhai, Liang Shen +11 · 8 citations
    Computer Science · #Advanced Neural Network Applications #Recommender Systems and Techniques #Stochastic Gradient Optimization Techniques
  2. Deep-FSMN for Large Vocabulary Continuous Speech Recognition
    2018/03/04 by Shiliang Zhang, Ming Lei, Zhang, Shiliang +5 · 6 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  3. SpeechLM: Enhanced Speech Pre-Training with Unpaired Textual Data
    2022/09/30 by Zhang, Ziqiang, Chen, Sanyuan, Zhou, Long +8 · 3 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  4. LDM-SVC: Latent Diffusion Model Based Zero-Shot Any-to-Any Singing Voice Conversion with Singer Guidance
    2024/06/08 by Shihao Chen, Yu Gu, Chen, Shihao +11 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Feedforward Sequential Memory Networks: A New Structure to Learn Long-term Dependency
    2015/12/28 by Shiliang Zhang, Cong Liu, Zhang, Shiliang +9 · 1 citation
    Computer Science · #FOS: Computer and information sciences #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
  6. Multi-Scale Attention with Dense Encoder for Handwritten Mathematical Expression Recognition
    2018/01/05 by Zhang, Jianshu, Du, Jun, Dai, Lirong · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. SiFiSinger: A High-Fidelity End-to-End Singing Voice Synthesizer based on Source-filter Model
    2024/10/16 by Cui, Jianwei, Gu, Yu, Weng, Chao +3 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  8. Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
    2022/03/31 by Ao, Junyi, Zhang, Ziqiang, Zhou, Long +7 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  9. Multichannel AV-wav2vec2: A Framework for Learning Multichannel Multi-Modal Speech Representation
    2024/01/07 by Zhu, Qiushi, Zhang, Jie, Gu, Yu +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  10. Deep CLAS: Deep Contextual Listen, Attend and Spell
    2024/09/26 by Mengzhi Wang, Wang, Mengzhi, Shifu Xiong +9 · 1 citation
    Health Professions · Computer Science · #Interpreting and Communication in Healthcare #Speech and dialogue systems #Natural Language Processing Techniques
  11. CSSinger: End-to-End Chunkwise Streaming Singing Voice Synthesis System Based on Conditional Variational Autoencoder
    2024/12/12 by Jianwei Cui, Yu Gu, Cui, Jianwei +9 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing