vix.ing · top · new · best · stats · spec

Ye Bai

  1. ADD 2022: the First Audio Deep Synthesis Detection Challenge
    2022/02/17 by Jiangyan Yi, Ruibo Fu, Yi, Jiangyan +36 · 22 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  2. Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
    2024/07/05 by Ye Bai, Jingping Chen, Bai, Ye +106 · 22 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
  3. Half-Truth: A Partially Fake Audio Detection Dataset
    2021/04/08 by Jiangyan Yi, Ye Bai, Yi, Jiangyan +12 · 6 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
  4. Fast End-to-End Speech Recognition via Non-Autoregressive Models and Cross-Modal Knowledge Transferring from BERT
    2021/02/15 by Ye Bai, Bai, Ye, Jiangyan Yi +9 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
  5. FSR: Accelerating the Inference Process of Transducer-Based Models by Applying Fast-Skip Regularization
    2021/04/07 by Zhengkun Tian, Tian, Zhengkun, Jiangyan Yi +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Learn Spelling from Teachers: Transferring Knowledge from Language Models to Sequence-to-Sequence Speech Recognition
    2019/07/13 by Ye Bai, Jiangyan Yi, Bai, Ye +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  7. Integrating Knowledge into End-to-End Speech Recognition from External Text-Only Data
    2019/12/04 by Ye Bai, Jiangyan Yi, Bai, Ye +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  8. Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition
    2020/05/16 by Zhengkun Tian, Jiangyan Yi, Tian, Zhengkun +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  9. Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition
    2020/05/11 by Ye Bai, Jiangyan Yi, Bai, Ye +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  10. Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
    2024/04/17 by Ye Bai, Bai, Ye, Chenxing Li +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering