Bai, Ye
- ADD 2022: the First Audio Deep Synthesis Detection Challenge
2022/02/17 by Jiangyan Yi, Ruibo Fu, Yi, Jiangyan +36 · 22 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024/07/05 by Ye Bai, Jingping Chen, Bai, Ye +106 · 22 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
- Half-Truth: A Partially Fake Audio Detection Dataset
2021/04/08 by Jiangyan Yi, Ye Bai, Yi, Jiangyan +12 · 6 citations
Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
- TraceableSpeech: Towards Proactively Traceable Text-to-Speech with Watermarking
2024/06/07 by Zhou, Junzuo, Yi, Jiangyan, Wang, Tao +5 · 7 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
2024/09/13 by Bai, Ye, Chen, Haonan, Chen, Jitong +35 · 6 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Fast End-to-End Speech Recognition via Non-Autoregressive Models and Cross-Modal Knowledge Transferring from BERT
2021/02/15 by Ye Bai, Bai, Ye, Jiangyan Yi +9 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- FSR: Accelerating the Inference Process of Transducer-Based Models by Applying Fast-Skip Regularization
2021/04/07 by Zhengkun Tian, Jiangyan Yi, Tian, Zhengkun +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Learn Spelling from Teachers: Transferring Knowledge from Language Models to Sequence-to-Sequence Speech Recognition
2019/07/13 by Ye Bai, Jiangyan Yi, Bai, Ye +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- Integrating Knowledge into End-to-End Speech Recognition from External Text-Only Data
2019/12/04 by Ye Bai, Bai, Ye, Jiangyan Yi +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition
2020/05/16 by Zhengkun Tian, Tian, Zhengkun, Jiangyan Yi +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition
2020/05/11 by Ye Bai, Jiangyan Yi, Bai, Ye +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- PolyVoice: Language Models for Speech to Speech Translation
2023/06/05 by Dong, Qianqian, Huang, Zhiying, Tian, Qiao +15 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
- Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
2024/04/17 by Ye Bai, Bai, Ye, Chenxing Li +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- P2Mark: Plug-and-play Parameter-level Watermarking for Neural Speech Generation
2025/04/07 by Ren, Yong, Yi, Jiangyan, Wang, Tao +7 · 1 citation
#FOS: Computer and information sciences #Sound (cs.SD)