vix.ing · top · new · best · stats · spec

Oneata, Dan

  1. Towards generalisable and calibrated synthetic speech detection with self-supervised representations
    2023/09/11 by Octavian Pascu, Adriana Stan, Pascu, Octavian +7 · 10 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Weakly-supervised deepfake localization in diffusion-generated images
    2023/11/08 by Dragos Tantaru, Elisabeta Oneaţă, Tantaru, Dragos +3 · 7 citations
    Computer Science · #Advanced Image Processing Techniques #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
  3. Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
    2024/11/29 by Stefan Smeu, Smeu, Stefan, Dragos-Alexandru Boldisor +7 · 1 voice · 9 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Machine Learning (cs.LG) #Sound (cs.SD) #cs.CV #cs.LG #cs.SD #eess.AS #eess.IV #electronic engineering #information engineering
  4. WavLM model ensemble for audio deepfake detection
    2024/08/14 by David Combei, Adriana Stan, Combei, David +5 · 9 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  5. DeCLIP: Decoding CLIP representations for deepfake localization
    2024/09/12 by Stefan Smeu, Elisabeta Oneaţă, Smeu, Stefan +3 · 7 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. Multilingual Multimodal Learning with Machine Translated Text
    2022/10/24 by Qiu, Chen, Oneata, Dan, Bugliarello, Emanuele +2 · 2 citations
    #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. YFACC: A Yorùbá speech-image dataset for cross-lingual keyword localisation through visual grounding
    2022/10/10 by Kayode Olaleye, Olaleye, Kayode, Dan Oneaţă +3 · 2 citations
    Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimodal Machine Learning Applications #Subtitles and Audiovisual Media #Video Analysis and Summarization #electronic engineering #information engineering
  8. Unmasking real-world audio deepfakes: A data-centric approach
    2025/06/11 by Combei, David, Stan, Adriana, Oneata, Dan +2 · 7 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  9. An evaluation of word-level confidence estimation for end-to-end automatic speech recognition
    2021/01/14 by Dan Oneaţă, Oneata, Dan, Alexandru Caranica +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  10. Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations
    2022/04/27 by Dan Oneaţă, Oneata, Dan, Horia Cucu +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Easy, Interpretable, Effective: openSMILE for voice deepfake detection
    2024/08/28 by Octavian Pascu, Pascu, Octavian, Dan Oneaţă +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis
  12. TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes
    2025/06/06 by Stan, Adriana, Combei, David, Oneata, Dan +1 · 6 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  13. Seeing What Tastes Good: Revisiting Multimodal Distributional Semantics in the Billion Parameter Era
    2025/06/04 by Dan Oneaţă, Dan Oneata, Oneata, Dan +4 · 1 voice · 2 citations
    Computer Science · Neuroscience · #Embodied and Extended Cognition #Face Recognition and Perception #Multimodal Machine Learning Applications #cs.CL #cs.CV