vix.ing · top · new · best · stats · spec

Bredin, Hervé

  1. pyannote.audio: neural building blocks for speaker diarization
    2019/11/04 by Bredin, Hervé, Yin, Ruiqing, Coria, Juan Manuel +7 · 26 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  2. End-to-end speaker segmentation for overlap-aware resegmentation
    2021/04/08 by Hervé Bredin, Bredin, Hervé, Antoine Laurent +1 · 7 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
  3. Overlap-aware low-latency online speaker diarization based on end-to-end\n local segmentation
    2021/09/14 by Juan Manuel Coria, Hervé Bredin, Coria, Juan M. +5 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  4. Overlap-aware diarization: resegmentation using neural end-to-end\n overlapped speech detection
    2019/10/25 by Latané Bullock, Hervé Bredin, Bullock, Latané +3 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Brouhaha: multi-task training for voice activity detection, speech-to-noise ratio, and C50 room acoustics estimation
    2022/10/24 by Marvin Lavechin, Lavechin, Marvin, Marianne Métais +17 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. The Speed Submission to DIHARD II: Contributions & Lessons Learned
    2019/11/06 by Sahidullah, Md, Patino, Jose, Cornell, Samuele +11 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  7. End-to-end Domain-Adversarial Voice Activity Detection
    2019/10/23 by Lavechin, Marvin, Gill, Marie-Philippe, Bousbib, Ruben +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #I.2.7 #electronic engineering #information engineering
  8. A Comparison of Metric Learning Loss Functions for End-To-End Speaker Verification
    2020/03/31 by Coria, Juan M., Bredin, Hervé, Ghannay, Sahar +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  9. TalTech-IRIT-LIS Speaker and Language Diarization Systems for DISPLACE 2024
    2024/07/17 by Joonas Kalda, Kalda, Joonas, Tanel Alumäe +9 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques