vix.ing · top · new · best · stats · spec

Najim Dehak

  1. x-vectors meet emotions: A study on dependencies between emotion and\n speaker recognition
    2020/02/12 by Raghavendra Pappagari, Tianzi Wang, Pappagari, Raghavendra +7 · 10 citations
    Computer Science · Psychology · #Speech Recognition and Synthesis #Emotion and Mood Recognition #Speech and Audio Processing
  2. ASSERT: Anti-Spoofing with Squeeze-Excitation and Residual neTworks
    2019/04/01 by Cheng-I Lai, Lai, Cheng-I, Nanxin Chen +5 · 8 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  3. WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis
    2021/06/17 by Nanxin Chen, Yu Zhang, Chen, Nanxin +11 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. DuTa-VC: A Duration-aware Typical-to-atypical Voice Conversion Approach with Diffusion Probabilistic Model
    2023/06/18 by Helin Wang, Thomas Thebaud, Wang, Helin +11 · 5 citations
    Computer Science · Health Professions · Medicine · #Audio and Speech Processing (eess.AS) #Dysphagia Assessment and Management #FOS: Electrical engineering #Signal Processing (eess.SP) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
  5. SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline
    2025/05/25 by Helin Wang, Wang, Helin, Jiarui Hai +17 · 1 voice · 4 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #cs.AI #cs.SD #eess.AS #electronic engineering #information engineering
  6. Study of Pre-processing Defenses against Adversarial Attacks on\n State-of-the-art Speaker Recognition Systems
    2021/01/21 by Sonal Joshi, Jesús Villalba, Joshi, Sonal +7 · 3 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Geophysical Methods and Applications
  7. Adversarial Attacks and Defenses for Speech Recognition Systems
    2021/03/31 by Piotr Żelasko, Sonal Joshi, Żelasko, Piotr +11 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Sound (cs.SD) #electronic engineering #information engineering
  8. SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer
    2024/09/12 by Helin Wang, Jiarui Hai, Wang, Helin +9 · 8 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
  9. DPM-TSE: A Diffusion Probabilistic Model for Target Sound Extraction
    2023/10/06 by Jiarui Hai, Helin Wang, Hai, Jiarui +9 · 4 citations
    Computer Science · Earth and Planetary Sciences · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #Underwater Acoustics Research #electronic engineering #information engineering
  10. Deep neural networks for emotion recognition combining audio and transcripts
    2019/11/01 by Jaejin Cho, Raghavendra Pappagari, Cho, Jaejin +9 · 3 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Electrical engineering #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
  11. Textual Data Augmentation for Arabic-English Code-Switching Speech Recognition
    2022/01/07 by Amir Hussein, Hussein, Amir, Shammur Absar Chowdhury +8 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  12. SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis
    2024/09/11 by Helin Wang, Wang, Helin, Yu Meng +13 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  13. Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
    2025/05/20 by Tiantian Feng, Jihwan Lee, Feng, Tiantian +21 · 11 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  14. Noise-robust Speech Separation with Fast Generative Correction
    2024/06/11 by Helin Wang, Jesús Villalba, Wang, Helin +9 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  15. Language model integration based on memory control for sequence to sequence speech recognition
    2018/11/06 by Jaejin Cho, Shinji Watanabe, Cho, Jaejin +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  16. MCE 2018: The 1st Multi-target Speaker Detection and Identification\n Challenge Evaluation
    2019/04/07 by Suwon Shon, Shon, Suwon, Najim Dehak +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  17. Punctuation Prediction in Spontaneous Conversations: Can We Mitigate ASR\n Errors with Retrofitted Word Embeddings?
    2020/04/13 by Łukasz Augustyniak, Augustyniak, Łukasz, Piotr Szymański +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  18. Learning Speaker Embedding from Text-to-Speech
    2020/10/21 by Jaejin Cho, Cho, Jaejin, Piotr Żelasko +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  19. Joint prediction of truecasing and punctuation for conversational speech\n in low-resource scenarios
    2021/09/13 by Raghavendra Pappagari, Piotr Żelasko, Pappagari, Raghavendra +7 · 1 citation
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems
  20. Representation Learning to Classify and Detect Adversarial Attacks\n against Speaker and Speech Recognition Systems
    2021/07/09 by Jesús Villalba, Sonal Joshi, Villalba, Jesús +5 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
  21. Discovering Phonetic Inventories with Crosslingual Automatic Speech Recognition
    2022/01/26 by Piotr Żelasko, Żelasko, Piotr, Siyuan Feng +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  22. AdvEst: Adversarial Perturbation Estimation to Classify and Detect Adversarial Attacks against Speaker Identification
    2022/04/08 by Sonal Joshi, Saurabh Kataria, Joshi, Sonal +5 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  23. Clean Label Attacks against SLU Systems
    2024/09/13 by Henry Li Xinyuan, Sonal Joshi, Xinyuan, Henry Li +9 · 2 citations
    Computer Science · #Advanced Authentication Protocols Security #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Security and Verification in Computing #Web Application Security Vulnerabilities #electronic engineering #information engineering
  24. CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
    2024/12/05 by Yen-Ju Lu, Lu, Yen-Ju, Jing Liu +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  25. Leveraging Gradient Reversal Loss and Multitask Learning for Datasets-Aware Audio Deepfake Detection
    2026/07/27 by Mingrui Liang, Thomas Thebaud, Lukasz Wojciak +4
    #eess.AS