vix.ing · top · new · best · stats · spec

Nirmesh J. Shah

  1. Quality assessment of voice converted speech using articulatory features
    2015/11/16 by Avni Rajpal, Rajpal, Avni, Nirmesh J. Shah +5 · 1 voice · 1 citation
    Computer Science · #FOS: Computer and information sciences #Sound (cs.SD) #cs.SD
  2. Nonparallel Emotional Voice Conversion For Unseen Speaker-Emotion Pairs Using Dual Domain Adversarial Network & Virtual Domain Pairing
    2023/02/21 by Nirmesh J. Shah, Shah, Nirmesh, Mayank Kumar Singh +5 · 5 citations
    Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing
    2024/06/13 by Neha Sahipjohn, Ashishkumar Gudmalwar, Sahipjohn, Neha +7 · 8 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. M2FNet: Multi-modal Fusion Network for Emotion Recognition in Conversation
    2022/06/05 by Vishal Chudasama, Chudasama, Vishal, Purbayan Kar +9 · 3 citations
    Psychology · Computer Science · #Emotion and Mood Recognition #Sentiment Analysis and Opinion Mining
  5. VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
    2024/06/12 by Ashishkumar Gudmalwar, Nirmesh J. Shah, Gudmalwar, Ashishkumar +7 · 3 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  6. Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning
    2024/03/20 by Shivam Mhaskar, Mhaskar, Shivam Ratnakant, Nirmesh J. Shah +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #electronic engineering #information engineering
  7. EmoReg: Directional Latent Vector Modeling for Emotional Intensity Regularization in Diffusion-based Voice Conversion
    2024/12/29 by Ashishkumar Gudmalwar, Ishan D. Biyani, Gudmalwar, Ashishkumar +7 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing