vix.ing · top · new · best · stats · spec

Dipjyoti Paul

  1. Enhancing Speech Intelligibility in Text-To-Speech Synthesis using\n Speaking Style Conversion
    2020/08/13 by Dipjyoti Paul, Paul, Dipjyoti, Muhammed PV Shifas +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  2. Speaker Conditional WaveRNN: Towards Universal Neural Vocoder for Unseen\n Speaker and Recording Conditions
    2020/08/09 by Dipjyoti Paul, Paul, Dipjyoti, Yannis Pantazis +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  3. Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers
    2026/07/26 by Dongseong Hwang, Prasanth Yadla, Kaan Elgin +8
    Computer Science · #cs.SD #cs.CL