vix.ing · top · new · best · stats · spec

ASMDD: Arabic Speech Mispronunciation Detection Dataset

2021/11/01 by Salah A. Aly, Aly, Salah A., Abdelrahman Salah +3
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Networking and Internet Architecture (cs.NI) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems

paper · pdf · doi:10.48550/arxiv.2111.01136

openalex publication_date 2021/11/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

The largest dataset of Arabic speech mispronunciation detections in Egyptian dialogues is introduced. The dataset is composed of annotated audio files representing the top 100 words that are most frequently used in the Arabic language, pronounced by 100 Egyptian children (aged between 2 and 8 years old). The dataset is collected and annotated on segmental pronunciation error detections by expert listeners.

Citations

Related