vix.ing · top · new · best · stats · spec

Stavros Petridis

  1. End-to-end Audiovisual Speech Recognition
    2018/02/18 by Stavros Petridis, Themos Stafylakis, Petridis, Stavros +9 · 1 voice · 4 citations
    #cs.CV
  2. Lips Don't Lie: A Generalisable and Robust Approach to Face Forgery\n Detection
    2020/12/14 by Alexandros Haliassos, Haliassos, Alexandros, Konstantinos Vougioukas +5 · 40 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
  3. Leveraging Real Talking Faces via Self-Supervision for Robust Forgery Detection
    2022/01/18 by Alexandros Haliassos, Haliassos, Alexandros, Rodrigo Mira +5 · 16 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis
  4. End-to-End Speech-Driven Facial Animation with Temporal GANs
    2018/05/23 by Konstantinos Vougioukas, Stavros Petridis, Vougioukas, Konstantinos +3 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Image and Video Processing (eess.IV) #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  5. Towards Pose-invariant Lip-Reading
    2019/11/14 by Shiyang Cheng, Pingchuan Ma, Cheng, Shiyang +11 · 2 citations
    Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Facial Nerve Paralysis Treatment and Research #Speech and Audio Processing
  6. Towards Practical Lipreading with Distilled and Efficient Models
    2020/07/13 by Pingchuan Ma, Brais Martínez, Ma, Pingchuan +5 · 2 citations
    Computer Science · Neuroscience · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Hand Gesture Recognition Systems #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Tactile and Sensory Interactions
  7. LA-VocE: Low-SNR Audio-visual Speech Enhancement using Neural Vocoders
    2022/11/20 by Rodrigo Mira, Buye Xu, Mira, Rodrigo +11 · 2 citations
    Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Machine Learning (cs.LG) #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  8. SynthVSR: Scaling Up Visual Speech Recognition With Synthetic Supervision
    2023/03/30 by Xubo Liu, Liu, Xubo, Egor Lakomkin +21 · 2 citations
    Computer Science · Engineering · #Speech and Audio Processing #Face recognition and analysis #Indoor and Outdoor Localization Technologies
  9. Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
    2024/11/04 by Alexandros Haliassos, Rodrigo Mira, Haliassos, Alexandros +9 · 4 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Speech and Audio Processing
  10. Contextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction
    2025/03/11 by Minsu Kim, Kim, Minsu, Rodrigo Mira +7 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  11. Hearing Loss Detection from Facial Expressions in One-on-one Conversations
    2024/01/17 by Yufeng Yin, Yin, Yufeng, Ishwarya Ananthabhotla +9 · 1 citation
    Computer Science · Neuroscience · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Hearing Loss and Rehabilitation #Speech and Audio Processing
  12. Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach
    2025/05/20 by Umberto Cappellazzo, Minsu Kim, Cappellazzo, Umberto +7 · 3 citations
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing