1999/10/13 by Anand Venkataraman, Venkataraman, Anand
Computer Science · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #I.2.6 #I.2.7 #Language Development and Disorders #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and dialogue systems #cs.CL #cs.LG
paper · pdf · doi:10.48550/arxiv.cs/9910011
48 pgs, 10 figs
arxiv created 1999/10/13 · openalex publication_date 1999/10/13 · arxiv updated 2009/11/30 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
A statistical model for segmentation and word discovery in child directed speech is presented. An incremental unsupervised learning algorithm to infer word boundaries based on this model is described and results of empirical tests showing that the algorithm is competitive with other models that have been used for similar tasks are also presented.