2026/02/14 by Edward Susko, Robert Lanfear, Andrew J Roger · 1 voice · 1 citation
Earth and Planetary Sciences · Biochemistry, Genetics and Molecular Biology · Mathematics · #Evolution and Paleontology Studies #Genomics and Phylogenetic Studies #Morphological variations and asymmetry
paper · doi:10.1093/sysbio/syag013
openalex created_date 2026/02/02 · openalex publication_date 2026/02/14 · openalex updated_date 2026/07/30
Sophisticated phylogenetic models often include mixture and/or partition model components. It was recently noted that information criteria tend to favour partition models over mixture models even in some cases where the latter are misspecified and give poor topological estimation. We show that this problem arises because partition models and mixture models fundamentally differ in their probability calculations: mixture models calculate likelihood contributions from sites as the marginal probability of the data averaging over parameter vectors that might have arisen at a site whereas partition model likelihood contributions are calculated as the probability of the site pattern conditional upon a fixed assigned parameter vector at that site. These differing probability calculations lead to AIC estimates that are not always comparable. We explore three generally applicable ways of correcting the issue.