2022/06/01 by Kavitha Raju, Raju, Kavitha, V. Anjaly +5
Computer Science · Arts and Humanities · #Speech Recognition and Synthesis #Music and Audio Processing #Diverse Musicological Studies
paper · pdf · doi:10.48550/arxiv.2206.01205
Automatic Speech Recognition (ASR) has increasing utility in the modern world. There are a many ASR models available for languages with large amounts of training data like English. However, low-resource languages are poorly represented. In response we create and release an open-licensed and formatted dataset of audio recordings of the Bible in low-resource northern Indian languages. We setup multiple experimental splits and train and analyze two competitive ASR models to serve as the baseline for future research using this data.