vix.ing · top · new · best · stats · spec

Cucu, Horia

  1. Adaptation of Whisper models to child speech recognition
    2023/07/24 by Rishabh Jain, Andrei Barcovschi, Jain, Rishabh +7 · 10 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  2. Towards generalisable and calibrated synthetic speech detection with self-supervised representations
    2023/09/11 by Octavian Pascu, Adriana Stan, Pascu, Octavian +7 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. A Wav2vec2-Based Experimental Study on Self-Supervised Learning Methods to Improve Child Speech Recognition
    2022/04/06 by Jain, Rishabh, Barcovschi, Andrei, Yiwere, Mariam +3 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  4. WavLM model ensemble for audio deepfake detection
    2024/08/14 by David Combei, Adriana Stan, Combei, David +5 · 9 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  5. Unmasking real-world audio deepfakes: A data-centric approach
    2025/06/11 by Combei, David, Stan, Adriana, Oneata, Dan +2 · 7 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  6. An evaluation of word-level confidence estimation for end-to-end automatic speech recognition
    2021/01/14 by Dan Oneaţă, Oneata, Dan, Alexandru Caranica +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Automated Circuit Sizing with Multi-objective Optimization based on Differential Evolution and Bayesian Inference
    2022/06/06 by Cătălin Vișan, Visan, Catalin, Octavian Pascu +13 · 1 citation
    Computer Science · Engineering · #Advanced Multi-Objective Optimization Algorithms #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #VLSI and FPGA Design Techniques
  8. Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations
    2022/04/27 by Dan Oneaţă, Oneata, Dan, Horia Cucu +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. A Text-to-Speech Pipeline, Evaluation Methodology, and Initial Fine-Tuning Results for Child Speech Synthesis
    2022/03/22 by Rishabh Jain, Mariam Yiwere, Jain, Rishabh +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  10. Easy, Interpretable, Effective: openSMILE for voice deepfake detection
    2024/08/28 by Octavian Pascu, Pascu, Octavian, Dan Oneaţă +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis
  11. TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes
    2025/06/06 by Stan, Adriana, Combei, David, Oneata, Dan +1 · 6 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering