vix.ing · top · new · best · stats · spec

Ao, Junyi

  1. SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
    2021/10/14 by Junyi Ao, Ao, Junyi, Rui Wang +25 · 24 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words
    2024/06/19 by Junyi Ao, Yuancheng Wang, Ao, Junyi +15 · 17 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  3. Multi-View Self-Attention Based Transformer for Speaker Recognition
    2021/10/11 by Rui Wang, Junyi Ao, Wang, Rui +13 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
  4. USED: Universal Speaker Extraction and Diarization
    2023/09/19 by Junyi Ao, Mehmet Sinan Yıldırım, Ao, Junyi +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT
    2022/03/29 by Rui Wang, Wang, Rui, Qibing Bai +15 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Pre-Training Transformer Decoder for End-to-End ASR Model with Unpaired Speech Data
    2022/03/31 by Ao, Junyi, Zhang, Ziqiang, Zhou, Long +7 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  7. CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning
    2022/10/08 by Chutong Meng, Junyi Ao, Meng, Chutong +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
  8. Overview of the Amphion Toolkit (v0.2)
    2025/01/26 by Jiaqi Li, Li, Jiaqi, Xueyao Zhang +20 · 5 citations
    Physics and Astronomy · Engineering · #Particle physics theoretical and experimental studies #Quantum Chromodynamics and Particle Interactions #Superconducting Materials and Applications
  9. EchoMind: An Interrelated Multi-level Benchmark for Evaluating Empathetic Speech Language Models
    2025/10/26 by Zhou, Li, Yu, Lutong, Lyu, You +6 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences