Bredin, Hervé
- pyannote.audio: neural building blocks for speaker diarization
2019/11/04 by Bredin, Hervé, Yin, Ruiqing, Coria, Juan Manuel +7 · 26 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- End-to-end speaker segmentation for overlap-aware resegmentation
2021/04/08 by Hervé Bredin, Bredin, Hervé, Antoine Laurent +1 · 7 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- Overlap-aware low-latency online speaker diarization based on end-to-end\n local segmentation
2021/09/14 by Juan Manuel Coria, Hervé Bredin, Coria, Juan M. +5 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Overlap-aware diarization: resegmentation using neural end-to-end\n overlapped speech detection
2019/10/25 by Latané Bullock, Hervé Bredin, Bullock, Latané +3 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Brouhaha: multi-task training for voice activity detection, speech-to-noise ratio, and C50 room acoustics estimation
2022/10/24 by Marvin Lavechin, Lavechin, Marvin, Marianne Métais +17 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- The Speed Submission to DIHARD II: Contributions & Lessons Learned
2019/11/06 by Sahidullah, Md, Patino, Jose, Cornell, Samuele +11 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- End-to-end Domain-Adversarial Voice Activity Detection
2019/10/23 by Lavechin, Marvin, Gill, Marie-Philippe, Bousbib, Ruben +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #I.2.7 #electronic engineering #information engineering
- A Comparison of Metric Learning Loss Functions for End-To-End Speaker Verification
2020/03/31 by Coria, Juan M., Bredin, Hervé, Ghannay, Sahar +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- TalTech-IRIT-LIS Speaker and Language Diarization Systems for DISPLACE 2024
2024/07/17 by Joonas Kalda, Kalda, Joonas, Tanel Alumäe +9 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques