Xiang Lv
- CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
2024/12/13 by Zhihao Du, Yuxuan Wang, Du, Zhihao +35 · 178 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
- CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
2025/05/23 by Zhihao Du, Du, Zhihao, Changfeng Gao +41 · 57 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- The USTC-Ximalaya system for the ICASSP 2022 multi-channel multi-party meeting transcription (M2MeT) challenge
2022/02/10 by Maokui He, Xiang Lv, He, Maokui +19 · 1 citation
Computer Science · Social Sciences · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Public Relations and Crisis Communication #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Qwen-Audio-3.0-TTS: Freely Controllable and Highly Robust Speech Synthesis with Multi-Stage Training Paradigm
2026/07/27 by Bajian Xiang, Cheng Wen, Han Zhao +12 · 1 citation
Engineering · #eess.AS