Zhengqi Wen
- ADD 2022: the First Audio Deep Synthesis Detection Challenge
2022/02/17 by Jiangyan Yi, Yi, Jiangyan, Ruibo Fu +36 · 23 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- ADD 2023: the Second Audio Deepfake Detection Challenge
2023/05/23 by Jiangyan Yi, Jianhua Tao, Yi, Jiangyan +33 · 25 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio
2024/05/08 by Yuankun Xie, Yi Lu, Xie, Yuankun +21 · 12 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Fast End-to-End Speech Recognition via Non-Autoregressive Models and Cross-Modal Knowledge Transferring from BERT
2021/02/15 by Ye Bai, Bai, Ye, Jiangyan Yi +9 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- FSR: Accelerating the Inference Process of Transducer-Based Models by Applying Fast-Skip Regularization
2021/04/07 by Zhengkun Tian, Tian, Zhengkun, Jiangyan Yi +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition
2020/05/16 by Zhengkun Tian, Tian, Zhengkun, Jiangyan Yi +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning
2025/05/21 by Jinyang Wu, Wu, Jinyang, Chan-Yu Liao +13 · 9 citations
Decision Sciences · #Complex Systems and Decision Making #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- VQ-CTAP: Cross-Modal Fine-Grained Sequence Representation Learning for Speech Processing
2024/08/11 by Chunyu Qiang, Qiang, Chunyu, Geng Wang +26 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Towards Fine-Grained Prosody Control for Voice Conversion
2019/10/24 by Zheng Lian, Zhengqi Wen, Lian, Zheng +1 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Learn Spelling from Teachers: Transferring Knowledge from Language Models to Sequence-to-Sequence Speech Recognition
2019/07/13 by Ye Bai, Bai, Ye, Jiangyan Yi +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- Integrating Knowledge into End-to-End Speech Recognition from External Text-Only Data
2019/12/04 by Ye Bai, Bai, Ye, Jiangyan Yi +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition
2020/05/11 by Ye Bai, Jiangyan Yi, Bai, Ye +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Generalized Source Tracing: Detecting Novel Audio Deepfake Algorithm with Real Emphasis and Fake Dispersion Strategy
2024/06/05 by Yuankun Xie, Ruibo Fu, Xie, Yuankun +13 · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Exploring the Role of Audio in Multimodal Misinformation Detection
2024/08/22 by Liu, Moyang, Yukun Liu, Liu, Yukun +10 · 2 citations
Social Sciences · #Misinformation and Its Impacts
- MDPE: A Multimodal Deception Dataset with Personality and Emotional Characteristics
2024/07/17 by Cong Cai, Shan Liang, Cai, Cong +25 · 2 citations
Computer Science · Psychology · Social Sciences · #Cybercrime and Law Enforcement Studies #Deception detection and forensic psychology #Crime Patterns and Interventions
- Learning From Yourself: A Self-Distillation Method for Fake Speech Detection
2023/03/02 by Cunhang Fan, Xue, Jun, Jiangyan Yi +10 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection
2025/05/16 by Hao Gu, Jiangyan Yi, Gu, Hao +15 · 3 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis #Speech Recognition and Synthesis #Music and Audio Processing
- DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
2025/01/29 by Mingkuan Feng, Feng, Mingkuan, Jinyang Wu +13 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- ASRRL-TTS: Agile Speaker Representation Reinforcement Learning for Text-to-Speech Speaker Adaptation
2024/07/07 by Ruibo Fu, Xin Qi, Fu, Ruibo +23 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
2024/08/20 by Yuankun Xie, Xie, Yuankun, Chenxu Xiong +21 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Signal Denoising Methods #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction
2025/05/31 by Cunhang Fan, Ying Chen, Fan, Cunhang +14 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- RadialRouter: Structured Representation for Efficient and Robust Large Language Models Routing
2025/06/04 by Ruihan Jin, Jin, Ruihan, Pengpeng Shao +11 · 3 citations
Computer Science · Physics and Astronomy · #Natural Language Processing Techniques #Complex Network Analysis Techniques #Advanced Graph Neural Networks
- SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
2026/07/16 by Jinyang Wu, Shuo Yang, Zhengxi Lu +8 · 1 voice
#cs.CL
- Decoupled Visual Processing: Efficient Multimodal Adaptation via Modality-Specific Transformer Substitution
2026/07/29 by Mingkuan Feng, Zhengqi Wen, Jianhua Tao
Computer Science · #cs.AI #cs.CV