vix.ing · top · new · best · stats · spec

Primitive Part-of-Speech Tagging using Word Length and Sentential Structure

1998/08/23 by Simon Cozens, Cozens, Simon
Computer Science · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Language Development and Disorders #Natural Language Processing Techniques #Speech and dialogue systems #cmp-lg #cs.CL

paper · pdf · doi:10.48550/arxiv.cmp-lg/9808011

6 pages

arxiv created 1998/08/23 · openalex publication_date 1998/08/23 · arxiv updated 2009/11/30 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

It has been argued that, when learning a first language, babies use a series of small clues to aid recognition and comprehension, and that one of these clues is word length. In this paper we present a statistical part of speech tagger which trains itself solely on the number of letters in each word in a sentence.

Related