1998/08/23 by Simon Cozens, Cozens, Simon
Computer Science · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Language Development and Disorders #Natural Language Processing Techniques #Speech and dialogue systems #cmp-lg #cs.CL
paper · pdf · doi:10.48550/arxiv.cmp-lg/9808011
6 pages
arxiv created 1998/08/23 · openalex publication_date 1998/08/23 · arxiv updated 2009/11/30 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
It has been argued that, when learning a first language, babies use a series of small clues to aid recognition and comprehension, and that one of these clues is word length. In this paper we present a statistical part of speech tagger which trains itself solely on the number of letters in each word in a sentence.