Najim Dehak
- x-vectors meet emotions: A study on dependencies between emotion and\n speaker recognition
2020/02/12 by Raghavendra Pappagari, Tianzi Wang, Pappagari, Raghavendra +7 · 10 citations
Computer Science · Psychology · #Speech Recognition and Synthesis #Emotion and Mood Recognition #Speech and Audio Processing
- ASSERT: Anti-Spoofing with Squeeze-Excitation and Residual neTworks
2019/04/01 by Cheng-I Lai, Lai, Cheng-I, Nanxin Chen +5 · 8 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
- WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis
2021/06/17 by Nanxin Chen, Yu Zhang, Chen, Nanxin +11 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- DuTa-VC: A Duration-aware Typical-to-atypical Voice Conversion Approach with Diffusion Probabilistic Model
2023/06/18 by Helin Wang, Thomas Thebaud, Wang, Helin +11 · 5 citations
Computer Science · Health Professions · Medicine · #Audio and Speech Processing (eess.AS) #Dysphagia Assessment and Management #FOS: Electrical engineering #Signal Processing (eess.SP) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline
2025/05/25 by Helin Wang, Wang, Helin, Jiarui Hai +17 · 1 voice · 4 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #cs.AI #cs.SD #eess.AS #electronic engineering #information engineering
- Study of Pre-processing Defenses against Adversarial Attacks on\n State-of-the-art Speaker Recognition Systems
2021/01/21 by Sonal Joshi, Jesús Villalba, Joshi, Sonal +7 · 3 citations
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Geophysical Methods and Applications
- Adversarial Attacks and Defenses for Speech Recognition Systems
2021/03/31 by Piotr Żelasko, Sonal Joshi, Żelasko, Piotr +11 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Sound (cs.SD) #electronic engineering #information engineering
- SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer
2024/09/12 by Helin Wang, Jiarui Hai, Wang, Helin +9 · 8 citations
Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
- DPM-TSE: A Diffusion Probabilistic Model for Target Sound Extraction
2023/10/06 by Jiarui Hai, Helin Wang, Hai, Jiarui +9 · 4 citations
Computer Science · Earth and Planetary Sciences · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #Underwater Acoustics Research #electronic engineering #information engineering
- Deep neural networks for emotion recognition combining audio and transcripts
2019/11/01 by Jaejin Cho, Raghavendra Pappagari, Cho, Jaejin +9 · 3 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Electrical engineering #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
- Textual Data Augmentation for Arabic-English Code-Switching Speech Recognition
2022/01/07 by Amir Hussein, Hussein, Amir, Shammur Absar Chowdhury +8 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
- SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis
2024/09/11 by Helin Wang, Wang, Helin, Yu Meng +13 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
2025/05/20 by Tiantian Feng, Jihwan Lee, Feng, Tiantian +21 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Noise-robust Speech Separation with Fast Generative Correction
2024/06/11 by Helin Wang, Jesús Villalba, Wang, Helin +9 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Language model integration based on memory control for sequence to sequence speech recognition
2018/11/06 by Jaejin Cho, Shinji Watanabe, Cho, Jaejin +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- MCE 2018: The 1st Multi-target Speaker Detection and Identification\n Challenge Evaluation
2019/04/07 by Suwon Shon, Shon, Suwon, Najim Dehak +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Punctuation Prediction in Spontaneous Conversations: Can We Mitigate ASR\n Errors with Retrofitted Word Embeddings?
2020/04/13 by Łukasz Augustyniak, Augustyniak, Łukasz, Piotr Szymański +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Learning Speaker Embedding from Text-to-Speech
2020/10/21 by Jaejin Cho, Cho, Jaejin, Piotr Żelasko +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Joint prediction of truecasing and punctuation for conversational speech\n in low-resource scenarios
2021/09/13 by Raghavendra Pappagari, Piotr Żelasko, Pappagari, Raghavendra +7 · 1 citation
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems
- Representation Learning to Classify and Detect Adversarial Attacks\n against Speaker and Speech Recognition Systems
2021/07/09 by Jesús Villalba, Sonal Joshi, Villalba, Jesús +5 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
- Discovering Phonetic Inventories with Crosslingual Automatic Speech Recognition
2022/01/26 by Piotr Żelasko, Żelasko, Piotr, Siyuan Feng +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- AdvEst: Adversarial Perturbation Estimation to Classify and Detect Adversarial Attacks against Speaker Identification
2022/04/08 by Sonal Joshi, Saurabh Kataria, Joshi, Sonal +5 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Clean Label Attacks against SLU Systems
2024/09/13 by Henry Li Xinyuan, Sonal Joshi, Xinyuan, Henry Li +9 · 2 citations
Computer Science · #Advanced Authentication Protocols Security #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Security and Verification in Computing #Web Application Security Vulnerabilities #electronic engineering #information engineering
- CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
2024/12/05 by Yen-Ju Lu, Lu, Yen-Ju, Jing Liu +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection
2026/07/27 by Mingrui Liang, Thomas Thebaud, Lukasz Wojciak +4
#eess.AS