Tseng, Yuan
- On the Utility of Self-supervised Models for Prosody-related Tasks
2022/10/13 by Guan-Ting Lin, Chi-Luen Feng, Lin, Guan-Ting +13 · 6 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems
2024/06/11 by Haibin Wu, Wu, Haibin, Yuan Tseng +3 · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- AV-SUPERB: A Multi-Task Evaluation Benchmark for Audio-Visual Representation Models
2023/09/19 by Tseng, Yuan, Berry, Layne, Chen, Yi-Ting +16 · 3 citations
#Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- CodecFake+: Codec-Based Resynthesized Data as a Proxy for Detecting CodecFake Speech
2025/01/14 by Xuanjun Chen, Jiawei Du, Chen, Xuanjun +19 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- REBORN: Reinforcement-Learned Boundary Segmentation with Iterative Training for Unsupervised ASR
2024/02/06 by Tseng, Liang-Hsuan, Hu, En-Pei, Chiang, Cheng-Han +4 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use
2025/05/27 by Parcollet, Titouan, Tseng, Yuan, Zhang, Shucong +1 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Evaluation of LLMs in Speech is Often Flawed: Test Set Contamination in Large Language Models for Speech Recognition
2025/05/28 by Yuan Tseng, Titouan Parcollet, Tseng, Yuan +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Authorship Attribution and Profiling #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering