S. Umesh
- DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utterances
2021/10/14 by Sreyan Ghosh, Ghosh, Sreyan, S Sakshi +6 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Sound (cs.SD) #electronic engineering #information engineering
- PADA: Pruning Assisted Domain Adaptation for Self-Supervised Speech Representations
2022/03/31 by Lodagala V S V Durga Prasad, Prasad, Lodagala V S V Durga, Sreyan Ghosh +3 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
- MAST: Multiscale Audio Spectrogram Transformers
2022/11/02 by Sreyan Ghosh, Ashish Seth, Ghosh, Sreyan +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- CCC-wav2vec 2.0: Clustering aided Cross Contrastive Self-supervised learning of speech representations
2022/10/05 by Vasista Sai Lodagala, Sreyan Ghosh, Lodagala, Vasista Sai +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
- FusDom: Combining In-Domain and Out-of-Domain Knowledge for Continuous Self-Supervised Learning
2023/12/20 by Ashish Seth, Seth, Ashish, Sreyan Ghosh +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Stable Distillation: Regularizing Continued Pre-training for Low-Resource Automatic Speech Recognition
2023/12/20 by Ashish Seth, Seth, Ashish, Sreyan Ghosh +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Improved cepstral mean and variance normalization using Bayesian framework
2013/12/01 by Nitin Prasad, S. Umesh · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing