Shah, Nirmesh
- Nonparallel Emotional Voice Conversion For Unseen Speaker-Emotion Pairs Using Dual Domain Adversarial Network & Virtual Domain Pairing
2023/02/21 by Shah, Nirmesh, Singh, Mayank Kumar, Takahashi, Naoya +1 · 3 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing
2024/06/13 by Neha Sahipjohn, Ashishkumar Gudmalwar, Sahipjohn, Neha +7 · 4 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- VECL-TTS: Voice identity and Emotional style controllable Cross-Lingual Text-to-Speech
2024/06/12 by Ashishkumar Gudmalwar, Gudmalwar, Ashishkumar, Nirmesh J. Shah +7 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- M2FNet: Multi-modal Fusion Network for Emotion Recognition in Conversation
2022/06/05 by Vishal Chudasama, Chudasama, Vishal, Purbayan Kar +9 · 1 citation
Psychology · Computer Science · #Emotion and Mood Recognition #Sentiment Analysis and Opinion Mining
- EmoReg: Directional Latent Vector Modeling for Emotional Intensity Regularization in Diffusion-based Voice Conversion
2024/12/29 by Ashishkumar Gudmalwar, Gudmalwar, Ashishkumar, Ishan D. Biyani +7 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing