2021/11/01 by Salah A. Aly, Aly, Salah A., Abdelrahman Salah +3
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Networking and Internet Architecture (cs.NI) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
paper · pdf · doi:10.48550/arxiv.2111.01136
openalex publication_date 2021/11/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
The largest dataset of Arabic speech mispronunciation detections in Egyptian dialogues is introduced. The dataset is composed of annotated audio files representing the top 100 words that are most frequently used in the Arabic language, pronounced by 100 Egyptian children (aged between 2 and 8 years old). The dataset is collected and annotated on segmental pronunciation error detections by expert listeners.