Kong Aik Lee
- ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech
2020/05/20 by Xin Wang, Junichi Yamagishi, Massimiliano Todisco +43 · 46 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
- ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
2024/08/16 by Xin Wang, Héctor Delgado, Wang, Xin +23 · 42 citations
Computer Science · #Hate Speech and Cyberbullying Detection
- Nes2Net: A Lightweight Nested Architecture for Foundation Model Driven Speech Anti-spoofing
2025/04/08 by Tianchi Liu, Liu, Tianchi, Duc-Tuan Truong +7 · 16 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning
2024/07/21 by Shuai Wang, Wang, Shuai, Zhengyang Chen +7 · 9 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- ASVspoof 2021: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan
2021/09/01 by Héctor Delgado, Delgado, Héctor, Nicholas Evans +19 · 4 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- Xi-Vector Embedding for Speaker Recognition
2021/08/12 by Kong Aik Lee, Lee, Kong Aik, Qiongqiong Wang +3 · 5 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
2022/02/08 by Fan Yu, Shiliang Zhang, Yu, Fan +29 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
2024/08/17 by Massimiliano Todisco, Michele Panariello, Todisco, Massimiliano +9 · 6 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
2025/02/13 by Xin Wang, Héctor Delgado, Wang, Xin +54 · 9 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Misinformation and Its Impacts #electronic engineering #information engineering
- LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
2024/09/23 by Hieu-Thi Luong, Haoyang Li, Luong, Hieu-Thi +7 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- The second multi-channel multi-party meeting transcription challenge (M2MeT) 2.0): A benchmark for speaker-attributed ASR
2023/09/24 by Yuhao Liang, Mohan Shi, Liang, Yuhao +24 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Self-supervised Speaker Recognition with Loss-gated Learning
2021/10/08 by Ruijie Tao, Tao, Ruijie, Kong Aik Lee +7 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Baseline Systems for the First Spoofing-Aware Speaker Verification Challenge: Score and Embedding Fusion
2022/04/21 by Hye-jin Shim, Shim, Hye-jin, Hemlata Tak +27 · 2 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- ASVspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech
2025/05/28 by Xin Wang, Héctor Delgado, Hemlata Tak +26 · 1 voice · 6 citations
Computer Science · #Speech Recognition and Synthesis #Network Security and Intrusion Detection #Hate Speech and Cyberbullying Detection
- Short-duration Speaker Verification (SdSV) Challenge 2021: the Challenge Evaluation Plan
2019/12/12 by Hossein Zeinali, Kong Aik Lee, Zeinali, Hossein +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Speaker recognition with two-step multi-modal deep cleansing
2022/10/28 by Ruijie Tao, Kong Aik Lee, Tao, Ruijie +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Towards Quantifying and Reducing Language Mismatch Effects in Cross-Lingual Speech Anti-Spoofing
2024/09/12 by Tianchi Liu, Ivan Kukanov, Liu, Tianchi +9 · 3 citations
Arts and Humanities · Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Freedom of Expression and Defamation #Hate Speech and Cyberbullying Detection #Linguistic research and analysis #Sound (cs.SD) #electronic engineering #information engineering
- Golden Gemini is All You Need: Finding the Sweet Spots for Speaker Verification
2023/12/06 by Tianchi Liu, Kong Aik Lee, Liu, Tianchi +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- An Empirical Bayes Framework for Open-Domain Dialogue Generation
2023/11/18 by Jing Yang Lee, Kong Aik Lee, Lee, Jing Yang +3 · 1 citation
Computer Science · #Topic Modeling #Speech and dialogue systems #Natural Language Processing Techniques
- VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
2024/03/01 by Weiwei Lin, Lin, Weiwei, Chenhang He +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Partially Randomizing Transformer Weights for Dialogue Response Diversity
2023/11/18 by Jing Yang Lee, Kong Aik Lee, Lee, Jing Yang +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
- On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
2024/09/15 by Junjie Li, Li, Junjie, Ke Zhang +9 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Advanced Data Compression Techniques
- MoMuSE: Momentum Multi-modal Target Speaker Extraction for Real-time Scenarios with Impaired Visual Cues
2024/12/11 by Junjie Li, Li, Junjie, Ke Zhang +7 · 2 citations
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- EDVD-LLaMA: Explainable Deepfake Video Detection via Multimodal Large Language Model Reasoning
2025/10/18 by Sun, Haoran, Chen Cai, Cai, Chen +8 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications
- EmoEUS: Uncertainty Supervision for Multimodal Emotion Recognition in Conversation
2026/07/19 by Zilong Huang, Kong Aik Lee, Junjie Li +2
#cs.MM #cs.CL #cs.LG
- EII-SCL: Harnessing Emotional Inertia for Multimodal Emotion Recognition in Conversation
2026/07/19 by Zilong Huang, Kong Aik Lee, Chong-Xin Gan +3
#cs.MM #cs.CL #cs.HC #eess.SP