vix.ing · top · new · best · stats · spec

Ravi Teja Gadde

  1. Diff2Lip: Audio Conditioned Diffusion Models for Lip-Synchronization
    2023/08/18 by Soumik Mukhopadhyay, Saksham Suri, Mukhopadhyay, Soumik +5 · 17 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Speech and Audio Processing
  2. Jasper: An End-to-End Convolutional Neural Acoustic Model
    2019/04/05 by Jason Li, Li, Jason, Vitaly Lavrukhin +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering