Nirmesh J. Shah
- Quality assessment of voice converted speech using articulatory features
2015/11/16 by Avni Rajpal, Rajpal, Avni, Nirmesh J. Shah +5 · 1 voice · 1 citation
Computer Science · #FOS: Computer and information sciences #Sound (cs.SD) #cs.SD
- Nonparallel Emotional Voice Conversion For Unseen Speaker-Emotion Pairs Using Dual Domain Adversarial Network & Virtual Domain Pairing
2023/02/21 by Nirmesh J. Shah, Shah, Nirmesh, Mayank Kumar Singh +5 · 5 citations
Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing
2024/06/13 by Neha Sahipjohn, Ashishkumar Gudmalwar, Sahipjohn, Neha +7 · 8 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- M2FNet: Multi-modal Fusion Network for Emotion Recognition in Conversation
2022/06/05 by Vishal Chudasama, Chudasama, Vishal, Purbayan Kar +9 · 3 citations
Psychology · Computer Science · #Emotion and Mood Recognition #Sentiment Analysis and Opinion Mining
- VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
2024/06/12 by Ashishkumar Gudmalwar, Nirmesh J. Shah, Gudmalwar, Ashishkumar +7 · 3 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning
2024/03/20 by Shivam Mhaskar, Mhaskar, Shivam Ratnakant, Nirmesh J. Shah +9 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #electronic engineering #information engineering
- EmoReg: Directional Latent Vector Modeling for Emotional Intensity Regularization in Diffusion-based Voice Conversion
2024/12/29 by Ashishkumar Gudmalwar, Ishan D. Biyani, Gudmalwar, Ashishkumar +7 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing