Odette Scharenborg
- Quantifying Bias in Automatic Speech Recognition
2021/03/28 by Siyuan Feng, Olya Kudina, Feng, Siyuan +5 · 4 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Predicting within and across language phoneme recognition performance of self-supervised learning speech pre-trained models
2022/06/24 by Hang Ji, Tanvina Patel, Ji, Hang +3 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- Exploring data augmentation in bias mitigation against non-native-accented speech
2023/12/24 by Yuanyuan Zhang, Zhang, Yuanyuan, Aaricia Herygers +7 · 6 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Phonetics and Phonology Research #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Self-supervised Speech Representations Still Struggle with African American Vernacular English
2024/08/26 by Kalvin Chang, Chang, Kalvin, Yi-Hui Chou +11 · 5 citations
Computer Science · #Speech Recognition and Synthesis
- AnyoneNet: Synchronized Speech and Talking Head Generation for Arbitrary Person
2021/08/09 by Xinsheng Wang, Qicong Xie, Wang, Xinsheng +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- That Sounds Familiar: an Analysis of Phonetic Representations Transfer\n Across Languages
2020/05/16 by Piotr Żelasko, Żelasko, Piotr, Laureano Moro-Velázquez +7 · 1 citation
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Manipulation of oral cancer speech using neural articulatory synthesis
2022/03/31 by Bence Mark Halpern, Teja Rebernik, Halpern, Bence Mark +13 · 2 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Voice and Speech Disorders #electronic engineering #information engineering
- S2IGAN: Speech-to-Image Generation via Adversarial Learning
2020/05/14 by Xinsheng Wang, Wang, Xinsheng, Tingting Qiao +7 · 1 citation
Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Music and Audio Processing
- Discovering Phonetic Inventories with Crosslingual Automatic Speech Recognition
2022/01/26 by Piotr Żelasko, Żelasko, Piotr, Siyuan Feng +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Good practices for evaluation of machine learning systems
2024/12/04 by Luciana Ferrer, Ferrer, Luciana, Odette Scharenborg +3 · 2 citations
Computer Science · Engineering · #Advanced Data Processing Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications
- The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition
2025/05/20 by Ming Gao, Gao, Ming, Shilong Wu +15 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Benchmarking Human and Automatic Speech Recognition of Diverse Speech: Initial Results
2026/07/21 by Ilse Huisman, Rares Popa, Yuanyuan Zhang +1
Computer Science · #cs.CL