Ao, Junyi
- SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
2021/10/14 by Junyi Ao, Ao, Junyi, Rui Wang +25 · 24 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words
2024/06/19 by Junyi Ao, Yuancheng Wang, Ao, Junyi +15 · 17 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
- Multi-View Self-Attention Based Transformer for Speaker Recognition
2021/10/11 by Rui Wang, Junyi Ao, Wang, Rui +13 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- USED: Universal Speaker Extraction and Diarization
2023/09/19 by Junyi Ao, Mehmet Sinan Yıldırım, Ao, Junyi +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT
2022/03/29 by Rui Wang, Wang, Rui, Qibing Bai +15 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
2022/03/31 by Ao, Junyi, Zhang, Ziqiang, Zhou, Long +7 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning
2022/10/08 by Chutong Meng, Junyi Ao, Meng, Chutong +7 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
- Overview of the Amphion Toolkit (v0.2)
2025/01/26 by Jiaqi Li, Li, Jiaqi, Xueyao Zhang +20 · 5 citations
Physics and Astronomy · Engineering · #Particle physics theoretical and experimental studies #Quantum Chromodynamics and Particle Interactions #Superconducting Materials and Applications
- EchoMind: An Interrelated Multi-level Benchmark for Evaluating Empathetic Speech Language Models
2025/10/26 by Zhou, Li, Yu, Lutong, Lyu, You +6 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences