Das, Rohan Kumar
- Voice Conversion Challenge 2020: Intra-lingual semi-parallel and cross-lingual voice conversion
2020/08/28 by Yi Zhao, Zhao, Yi, Wen-Chin Huang +13 · 7 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Topic Modeling #Voice and Speech Disorders
- XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection
2024/11/15 by Yang Xiao, Xiao, Yang, Rohan Kumar Das +1 · 17 citations
Computer Science · #Advanced Malware Detection Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Network Security and Intrusion Detection #Sound (cs.SD) #Spam and Phishing Detection #electronic engineering #information engineering
- Light Convolutional Neural Network with Feature Genuinization for Detection of Synthetic Speech Attacks
2020/09/21 by Zhenzong Wu, Rohan Kumar Das, Wu, Zhenzong +5 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Nes2Net: A Lightweight Nested Architecture for Foundation Model Driven Speech Anti-spoofing
2025/04/08 by Tianchi Liu, Duc-Tuan Truong, Liu, Tianchi +7 · 16 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Exploring Text-Queried Sound Event Detection with Audio Source Separation
2024/09/20 by Han Yin, Yin, Han, Jisheng Bai +15 · 8 citations
Computer Science · #Advanced Text Analysis Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Audio-visual Speaker Recognition with a Cross-modal Discriminative Network
2020/08/10 by Tao, Ruijie, Das, Rohan Kumar, Li, Haizhou · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
- MFA: TDNN with Multi-scale Frequency-channel Attention for Text-independent Speaker Verification with Short Utterances
2022/02/03 by Liu, Tianchi, Das, Rohan Kumar, Lee, Kong Aik +1 · 3 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Signal Processing (eess.SP) #Sound (cs.SD) #electronic engineering #information engineering
- How Do Neural Spoofing Countermeasures Detect Partially Spoofed Audio?
2024/06/04 by Tianchi Liu, Liu, Tianchi, Lin Zhang +9 · 5 citations
Computer Science · Neuroscience · #Speech and Audio Processing #Music and Audio Processing #Hearing Loss and Rehabilitation
- TF-Mamba: A Time-Frequency Network for Sound Source Localization
2024/09/08 by Yang Xiao, Xiao, Yang, Rohan Kumar Das +1 · 4 citations
Computer Science · #Speech and Audio Processing #Music and Audio Processing #Music Technology and Sound Studies
- Self-supervised Speaker Recognition with Loss-gated Learning
2021/10/08 by Ruijie Tao, Kong Aik Lee, Tao, Ruijie +7 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- EnvSDD: Benchmarking Environmental Sound Deepfake Detection
2025/05/25 by Yin, Han, Xiao, Yang, Das, Rohan Kumar +4 · 8 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Black-box Attacks on Automatic Speaker Verification using Feedback-controlled Voice Conversion
2019/09/17 by Xiaohai Tian, Tian, Xiaohai, Rohan Kumar Das +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
2024/07/04 by Yang Xiao, Rohan Kumar Das, Xiao, Yang +1 · 4 citations
Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- The Attacker's Perspective on Automatic Speaker Verification: An Overview
2020/04/19 by Das, Rohan Kumar, Tian, Xiaohai, Kinnunen, Tomi +1 · 1 citation
#Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
- The INTERSPEECH 2020 Far-Field Speaker Verification Challenge
2020/05/16 by Xiaoyi Qin, Ming Li, Qin, Xiaoyi +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Predictions of Subjective Ratings and Spoofing Assessments of Voice Conversion Challenge 2020 Submissions
2020/09/08 by Rohan Kumar Das, Das, Rohan Kumar, Tomi Kinnunen +13 · 1 citation
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
- WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
2024/07/04 by Yang Xiao, Xiao, Yang, Rohan Kumar Das +1 · 3 citations
Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- Significance of Data Augmentation for Improving Cleft Lip and Palate Speech Recognition
2021/10/02 by Sudro, Protima Nomo, Das, Rohan Kumar, Sinha, Rohit +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Graph Fourier Transform based Audio Zero-watermarking
2021/09/16 by Xu, Longting, Huang, Daiyu, Zaidi, Syed Faham Ali +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- A Multi-Task Learning Framework for Sound Event Detection using High-level Acoustic Characteristics of Sounds
2023/05/18 by Tanmay Khandelwal, Rohan Kumar Das, Khandelwal, Tanmay +1 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
- Leveraging Audio-Tagging Assisted Sound Event Detection using Weakified Strong Labels and Frequency Dynamic Convolutions
2023/04/25 by Tanmay Khandelwal, Rohan Kumar Das, Khandelwal, Tanmay +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
- ESDD 2026: Environmental Sound Deepfake Detection Challenge Evaluation Plan
2025/08/06 by Yin, Han, Xiao, Yang, Das, Rohan Kumar +2 · 4 citations
#FOS: Computer and information sciences #Sound (cs.SD)
- Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan
2025/08/06 by Moscati, Marta, Abdullah, Ahmed, Saeed, Muhammad Saad +7 · 5 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- DG-SED: Domain Generalization for Sound Event Detection with Heterogeneous Training Data
2024/07/04 by Yang Xiao, Han Yin, Xiao, Yang +5 · 2 citations
Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
- Where's That Voice Coming? Continual Learning for Sound Source Localization
2024/07/04 by Yang Xiao, Rohan Kumar Das, Xiao, Yang +1 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Neural Networks and Applications #Sound (cs.SD) #electronic engineering #information engineering
- RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing
2025/07/11 by Xiao Yang, Xiao, Yang, Ting Dang +3 · 2 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan
2024/04/14 by Muhammad Saad Saeed, Shah Nawaz, Saeed, Muhammad Saad +17 · 2 citations
Arts and Humanities · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Subtitles and Audiovisual Media #electronic engineering #information engineering
- AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
2025/05/17 by Yang Xiao, Tianyi Peng, Xiao, Yang +7 · 3 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Domain Adaptation and Few-Shot Learning #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Speaker-Utterance Dual Attention for Speaker and Utterance Verification
2020/08/20 by Tianchi Liu, Liu, Tianchi, Rohan Kumar Das +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Listen, Analyze, and Adapt to Learn New Attacks: An Exemplar-Free Class Incremental Learning Method for Audio Deepfake Source Tracing
2025/05/20 by Xiao, Yang, Das, Rohan Kumar · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering