Stavros Petridis
- End-to-end Audiovisual Speech Recognition
2018/02/18 by Stavros Petridis, Themos Stafylakis, Petridis, Stavros +9 · 1 voice · 4 citations
#cs.CV
- Lips Don't Lie: A Generalisable and Robust Approach to Face Forgery\n Detection
2020/12/14 by Alexandros Haliassos, Haliassos, Alexandros, Konstantinos Vougioukas +5 · 40 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
- Leveraging Real Talking Faces via Self-Supervision for Robust Forgery Detection
2022/01/18 by Alexandros Haliassos, Haliassos, Alexandros, Rodrigo Mira +5 · 16 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis
- End-to-End Speech-Driven Facial Animation with Temporal GANs
2018/05/23 by Konstantinos Vougioukas, Stavros Petridis, Vougioukas, Konstantinos +3 · 6 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Image and Video Processing (eess.IV) #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Towards Pose-invariant Lip-Reading
2019/11/14 by Shiyang Cheng, Pingchuan Ma, Cheng, Shiyang +11 · 2 citations
Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Facial Nerve Paralysis Treatment and Research #Speech and Audio Processing
- Towards Practical Lipreading with Distilled and Efficient Models
2020/07/13 by Pingchuan Ma, Brais Martínez, Ma, Pingchuan +5 · 2 citations
Computer Science · Neuroscience · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Hand Gesture Recognition Systems #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Tactile and Sensory Interactions
- LA-VocE: Low-SNR Audio-visual Speech Enhancement using Neural Vocoders
2022/11/20 by Rodrigo Mira, Buye Xu, Mira, Rodrigo +11 · 2 citations
Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Machine Learning (cs.LG) #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- SynthVSR: Scaling Up Visual Speech Recognition With Synthetic Supervision
2023/03/30 by Xubo Liu, Liu, Xubo, Egor Lakomkin +21 · 2 citations
Computer Science · Engineering · #Speech and Audio Processing #Face recognition and analysis #Indoor and Outdoor Localization Technologies
- Unified Speech Recognition: A Single Model for Auditory, Visual, and Audiovisual Inputs
2024/11/04 by Alexandros Haliassos, Rodrigo Mira, Haliassos, Alexandros +9 · 4 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Speech and Audio Processing
- Contextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction
2025/03/11 by Minsu Kim, Kim, Minsu, Rodrigo Mira +7 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Hearing Loss Detection from Facial Expressions in One-on-one Conversations
2024/01/17 by Yufeng Yin, Yin, Yufeng, Ishwarya Ananthabhotla +9 · 1 citation
Computer Science · Neuroscience · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Hearing Loss and Rehabilitation #Speech and Audio Processing
- Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach
2025/05/20 by Umberto Cappellazzo, Minsu Kim, Cappellazzo, Umberto +7 · 3 citations
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing