Huang, Chien-yu
- Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
2023/09/18 by Chien‐Yu Huang, Ke-Han Lu, Huang, Chien-yu +27 · 20 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
2024/11/08 by Huang, Chien-yu, Chen, Wei-Chih, Yang, Shu-wen +77 · 27 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
- Defending Your Voice: Adversarial Attack on Voice Conversion
2020/05/18 by Huang, Chien-yu, Lin, Yist Y., Lee, Hung-yi +1 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- How Far Are We from Robust Voice Conversion: A Survey
2020/11/24 by Huang, Tzu-hsien, Lin, Jheng-hao, Huang, Chien-yu +1 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
2025/07/03 by Lu, Ke-Han, Chen, Zhehuai, Fu, Szu-Wei +25 · 19 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Investigating on Incorporating Pretrained and Learnable Speaker\n Representations for Multi-Speaker Multi-Style Text-to-Speech
2021/03/06 by Chung-Ming Chien, Jheng-Hao Lin, Chien, Chung-Ming +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Toward Degradation-Robust Voice Conversion
2021/10/14 by Chien‐Yu Huang, Huang, Chien-yu, Kai-Wei Chang +3 · 2 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- A Preliminary Exploration with GPT-4o Voice Mode
2025/02/14 by Yuxiang Lin, Chih-Kai Yang, Lin, Yu-Xiang +11 · 5 citations
Computer Science · #Computational Physics and Python Applications
- Utilizing Self-supervised Representations for MOS Prediction
2021/04/07 by Tseng, Wei-Cheng, Huang, Chien-yu, Kao, Wei-Tsung +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering