Dipjyoti Paul
- Enhancing Speech Intelligibility in Text-To-Speech Synthesis using\n Speaking Style Conversion
2020/08/13 by Dipjyoti Paul, Paul, Dipjyoti, Muhammed PV Shifas +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Speaker Conditional WaveRNN: Towards Universal Neural Vocoder for Unseen\n Speaker and Recording Conditions
2020/08/09 by Dipjyoti Paul, Paul, Dipjyoti, Yannis Pantazis +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers
2026/07/26 by Dongseong Hwang, Prasanth Yadla, Kaan Elgin +8
Computer Science · #cs.SD #cs.CL