1999/11/01 by Don Ringe · 2 citations
Arts and Humanities · Computer Science · Mathematics · #Lexicography and Language Studies #Natural Language Processing Techniques #Linguistics and language evolution #Simple (philosophy) #Computer science #Sample (material) #Root (linguistics) #Cognate #Test (biology) #Algorithm #Mathematics #Statistics #Artificial intelligence #Linguistics #Biology #Philosophy #Epistemology
paper · doi:10.1111/1467-968x.00049
openalex publication_date 1999/11/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/05/21
Straightforward application of the ‘same birthdays’ problem shows that it is surprisingly easy to find a match for any CVC‐root when wordlists of many languages are compared simultaneously. A simple algorithm can be used to estimate the expected incidence of multiple CVC‐matches. A test of a sample from Greenberg's ‘Amerind Etymological Dictionary’ using that algorithm finds no ‘cognate sets’ whose resemblances are clearly not the result of random factors. The same test can and should be applied to all comparative etymological dictionaries.