vix.ing · top · new · best · stats · spec

Huang, Chien-yu

  1. Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
    2023/09/18 by Chien‐Yu Huang, Ke-Han Lu, Huang, Chien-yu +27 · 20 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
    2024/11/08 by Huang, Chien-yu, Chen, Wei-Chih, Yang, Shu-wen +77 · 27 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  3. Defending Your Voice: Adversarial Attack on Voice Conversion
    2020/05/18 by Huang, Chien-yu, Lin, Yist Y., Lee, Hung-yi +1 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. How Far Are We from Robust Voice Conversion: A Survey
    2020/11/24 by Huang, Tzu-hsien, Lin, Jheng-hao, Huang, Chien-yu +1 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  5. DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
    2025/07/03 by Lu, Ke-Han, Chen, Zhehuai, Fu, Szu-Wei +25 · 19 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  6. Investigating on Incorporating Pretrained and Learnable Speaker\n Representations for Multi-Speaker Multi-Style Text-to-Speech
    2021/03/06 by Chung-Ming Chien, Jheng-Hao Lin, Chien, Chung-Ming +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Toward Degradation-Robust Voice Conversion
    2021/10/14 by Chien‐Yu Huang, Huang, Chien-yu, Kai-Wei Chang +3 · 2 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
  8. A Preliminary Exploration with GPT-4o Voice Mode
    2025/02/14 by Yuxiang Lin, Chih-Kai Yang, Lin, Yu-Xiang +11 · 5 citations
    Computer Science · #Computational Physics and Python Applications
  9. Utilizing Self-supervised Representations for MOS Prediction
    2021/04/07 by Tseng, Wei-Cheng, Huang, Chien-yu, Kao, Wei-Tsung +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering