2016/01/11 by Claudia Peersman, Walter Daelemans, Peersman, Claudia +7
Computer Science · Social Sciences · #Authorship Attribution and Profiling #Computation and Language (cs.CL) #Digital Communication and Language #FOS: Computer and information sciences #Linguistic Variation and Morphology #cs.CL
paper · pdf · doi:10.48550/arxiv.1601.02431
arxiv created 2016/01/11 · openalex publication_date 2016/01/11 · arxiv updated 2016/01/12 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
We present a corpus-based analysis of the effects of age, gender and region of origin on the production of both "netspeak" or "chatspeak" features and regional speech features in Flemish Dutch posts that were collected from a Belgian online social network platform. The present study shows that combining quantitative and qualitative approaches is essential for understanding non-standard linguistic variation in a CMC corpus. It also presents a methodology that enables the systematic study of this variation by including all non-standard words in the corpus. The analyses resulted in a convincing illustration of the Adolescent Peak Principle. In addition, our approach revealed an intriguing correlation between the use of regional speech features and chatspeak features.