vix.ing · top · new · best · stats · spec

S. Umesh

  1. DeToxy: A Large-Scale Multimodal Dataset for Toxicity Classification in Spoken Utterances
    2021/10/14 by Sreyan Ghosh, Ghosh, Sreyan, S Sakshi +6 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Hate Speech and Cyberbullying Detection #Sound (cs.SD) #electronic engineering #information engineering
  2. PADA: Pruning Assisted Domain Adaptation for Self-Supervised Speech Representations
    2022/03/31 by Lodagala V S V Durga Prasad, Prasad, Lodagala V S V Durga, Sreyan Ghosh +3 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
  3. MAST: Multiscale Audio Spectrogram Transformers
    2022/11/02 by Sreyan Ghosh, Ashish Seth, Ghosh, Sreyan +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. CCC-wav2vec 2.0: Clustering aided Cross Contrastive Self-supervised learning of speech representations
    2022/10/05 by Vasista Sai Lodagala, Sreyan Ghosh, Lodagala, Vasista Sai +3 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
  5. FusDom: Combining In-Domain and Out-of-Domain Knowledge for Continuous Self-Supervised Learning
    2023/12/20 by Ashish Seth, Seth, Ashish, Sreyan Ghosh +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  6. Stable Distillation: Regularizing Continued Pre-training for Low-Resource Automatic Speech Recognition
    2023/12/20 by Ashish Seth, Seth, Ashish, Sreyan Ghosh +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Improved cepstral mean and variance normalization using Bayesian framework
    2013/12/01 by Nitin Prasad, S. Umesh · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing