Cong Han
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
2023/06/13 by Yinghao Aaron Li, Li, Yinghao Aaron, Cong Han +7 · 1 voice · 30 citations
#eess.AS #cs.AI #cs.CL #cs.LG #cs.SD
- FaSNet: Low-latency Adaptive Beamforming for Multi-microphone Audio Processing
2019/09/29 by Yi Luo, Luo, Yi, Enea Ceolini +7 · 10 citations
Computer Science · Engineering · #Speech and Audio Processing #Music and Audio Processing #Advanced Adaptive Filtering Techniques
- Improving Conversational Recommendation Systems' Quality with Context-Aware Item Meta Information
2021/12/15 by Bowen Yang, Cong Han, Yang, Bowen +7 · 4 citations
Computer Science · Psychology · #Topic Modeling #Recommender Systems and Techniques #Mental Health via Writing
- StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis
2022/05/30 by Yinghao Aaron Li, Cong Han, Li, Yinghao Aaron +3 · 4 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Listen, Chat, and Remix: Text-Guided Soundscape Remixing for Enhanced Auditory Experience
2024/02/06 by Xilin Jiang, Cong Han, Jiang, Xilin +5 · 3 citations
Health Professions · #Noise Effects and Management
- StyleTTS-VC: One-Shot Voice Conversion by Knowledge Transfer from Style-Based TTS Models
2022/12/29 by Yinghao Aaron Li, Cong Han, Li, Yinghao Aaron +3 · 2 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- HiFTNet: A Fast High-Quality Neural Vocoder with Harmonic-plus-Noise Filter and Inverse Short Time Fourier Transform
2023/09/18 by Yinghao Aaron Li, Cong Han, Li, Yinghao Aaron +5 · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Exploring Self-Supervised Contrastive Learning of Spatial Sound Event Representation
2023/09/27 by Xilin Jiang, Cong Han, Jiang, Xilin +5 · 2 citations
Computer Science · Engineering · #Music and Audio Processing #Speech and Audio Processing #Acoustic Wave Phenomena Research
- Online Binaural Speech Separation of Moving Speakers With a Wavesplit Network
2023/03/13 by Cong Han, Nima Mesgarani, Han, Cong +1 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- d-Alanine:d-alanine ligase as a new target for the flavonoids quercetin and apigenin
2008/09/07 by Dalei Wu, Yunhua Kong, Cong Han +4 · 25 citations
Biochemistry, Genetics and Molecular Biology · Medicine · #Bioactive natural compounds #Synthesis of Organic Compounds #Phytochemicals and Antioxidant Activities
- Continuous Speech Separation Using Speaker Inventory for Long Multi-talker Recording
2020/12/17 by Cong Han, Yi Luo, Han, Cong +19 · 1 citation
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing