vix.ing · top · new · best · stats · spec

Horia Cucu

  1. Adaptation of Whisper models to child speech recognition
    2023/07/24 by Rishabh Jain, Andrei Barcovschi, Jain, Rishabh +7 · 10 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  2. Towards generalisable and calibrated synthetic speech detection with self-supervised representations
    2023/09/11 by Octavian Pascu, Pascu, Octavian, Adriana Stan +7 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. WavLM model ensemble for audio deepfake detection
    2024/08/14 by David Combei, Combei, David, Adriana Stan +5 · 9 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  4. An evaluation of word-level confidence estimation for end-to-end automatic speech recognition
    2021/01/14 by Dan Oneaţă, Alexandru Caranica, Oneata, Dan +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  5. Automated Circuit Sizing with Multi-objective Optimization based on Differential Evolution and Bayesian Inference
    2022/06/06 by Cătălin Vișan, Visan, Catalin, Octavian Pascu +13 · 1 citation
    Computer Science · Engineering · #Advanced Multi-Objective Optimization Algorithms #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #VLSI and FPGA Design Techniques
  6. Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations
    2022/04/27 by Dan Oneaţă, Oneata, Dan, Horia Cucu +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. A Text-to-Speech Pipeline, Evaluation Methodology, and Initial Fine-Tuning Results for Child Speech Synthesis
    2022/03/22 by Rishabh Jain, Mariam Yiwere, Jain, Rishabh +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  8. Easy, Interpretable, Effective: openSMILE for voice deepfake detection
    2024/08/28 by Octavian Pascu, Pascu, Octavian, Dan Oneaţă +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis