vix.ing · top · new · best · stats · spec

How hard is it to match CVC‐roots?

1999/11/01 by Don Ringe · 2 citations
Arts and Humanities · Computer Science · Mathematics · #Lexicography and Language Studies #Natural Language Processing Techniques #Linguistics and language evolution #Simple (philosophy) #Computer science #Sample (material) #Root (linguistics) #Cognate #Test (biology) #Algorithm #Mathematics #Statistics #Artificial intelligence #Linguistics #Biology #Philosophy #Epistemology

paper · doi:10.1111/1467-968x.00049

openalex publication_date 1999/11/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/05/21

Abstract

Straightforward application of the ‘same birthdays’ problem shows that it is surprisingly easy to find a match for any CVC‐root when wordlists of many languages are compared simultaneously. A simple algorithm can be used to estimate the expected incidence of multiple CVC‐matches. A test of a sample from Greenberg's ‘Amerind Etymological Dictionary’ using that algorithm finds no ‘cognate sets’ whose resemblances are clearly not the result of random factors. The same test can and should be applied to all comparative etymological dictionaries.

Cited by