vix.ing · top · new · best · stats · spec

Lei, Yi

  1. MsEmoTTS: Multi-scale emotion transfer, prediction, and control for emotional speech synthesis
    2022/01/17 by Yi Lei, Shan Yang, Lei, Yi +5 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  2. METTS: Multilingual Emotional Text-to-Speech by Cross-speaker and Cross-lingual Emotion Transfer
    2023/07/29 by Xinfa Zhu, Yi Lei, Zhu, Xinfa +11 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. DSPGAN: a GAN-based universal vocoder for high-fidelity TTS by time-frequency domain supervision from DSP
    2022/11/02 by Kun Song, Yongmao Zhang, Song, Kun +13 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. PromptStyle: Controllable Style Transfer for Text-to-Speech with Natural Language Descriptions
    2023/05/31 by Yongmao Zhang, Liu, Guanghou, Yi Lei +10 · 5 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  5. PromptVC: Flexible Stylistic Voice Conversion in Latent Space Driven by Natural Language Prompts
    2023/09/17 by Jixun Yao, Yuguang Yang, Yao, Jixun +17 · 5 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
  6. Fine-grained Emotion Strength Transfer, Control and Prediction for Emotional Speech Synthesis
    2020/11/17 by Yi Lei, Lei, Yi, Shan Yang +3 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Vec-Tok Speech: speech vectorization and tokenization for neural speech generation
    2023/10/11 by Xinfa Zhu, Yuanjun Lv, Zhu, Xinfa +13 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  8. Distinguishable Speaker Anonymization based on Formant and Fundamental Frequency Scaling
    2022/11/06 by Yao, Jixun, Wang, Qing, Lei, Yi +4 · 3 citations
    #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  9. Preserving background sound in noise-robust voice conversion via multi-task learning
    2022/11/06 by Jixun Yao, Yi Lei, Yao, Jixun +15 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. PromptSpeaker: Speaker Generation Based on Text Descriptions
    2023/10/08 by Zhang, Yongmao, Liu, Guanghou, Lei, Yi +4 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. UniSyn: An End-to-End Unified Model for Text-to-Speech and Singing Voice Synthesis
    2022/12/03 by Yi Lei, Lei, Yi, Shan Yang +11 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  12. Multi-Speaker Expressive Speech Synthesis via Multiple Factors Decoupling
    2022/11/19 by Xinfa Zhu, Yi Lei, Zhu, Xinfa +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  13. VITS-based Singing Voice Conversion System with DSPGAN post-processing for SVCC2023
    2023/10/08 by Yiquan Zhou, Meng Chen, Zhou, Yiquan +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Zero-Shot Emotion Transfer For Cross-Lingual Speech Synthesis
    2023/10/06 by Yuke Li, Xinfa Zhu, Li, Yuke +11 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  15. StyleS2ST: Zero-shot Style Transfer for Direct Speech-to-speech Translation
    2023/05/28 by Kun Song, Song, Kun, Yi Ren +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  16. Glow-WaveGAN 2: High-quality Zero-shot Text-to-speech Synthesis and Any-to-any Voice Conversion
    2022/07/05 by Lei, Yi, Yang, Shan, Cong, Jian +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  17. Accent-VITS:accent transfer for end-to-end TTS
    2023/12/28 by Linhan Ma, Ma, Linhan, Yongmao Zhang +11 · 1 citation
    Computer Science · Psychology · Medicine · #Speech Recognition and Synthesis #Phonetics and Phonology Research #Voice and Speech Disorders
  18. DenseTrack: Drone-based Crowd Tracking via Density-aware Motion-appearance Synergy
    2024/07/24 by Lei, Yi, Zhu, Huilin, Yuan, Jingling +3 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences