2012/01/06 by Mark Tygert, Tygert, Mark
Chemistry · Mathematics · Social Sciences · #Census and Population Estimation #Computation (stat.CO) #Data Analysis and Archiving #FOS: Computer and information sciences #History and advancements in chemistry #Methodology (stat.ME) #stat.CO #stat.ME
paper · pdf · doi:10.48550/arxiv.1201.1421
14 pages, 18 tables
arxiv created 2012/01/06 · openalex publication_date 2012/01/06 · arxiv updated 2012/01/09 · openalex created_date 2016/06/24 · openalex updated_date 2026/07/28
The model for homogeneity of proportions in a two-way contingency-table/cross-tabulation is the same as the model of independence, except that the probabilistic process generating the data is viewed as fixing the column totals (but not the row totals). When gauging the consistency of observed data with the assumption of independence, recent work has illustrated that the Euclidean/Frobenius/Hilbert-Schmidt distance is often far more statistically powerful than the classical statistics such as chi-square, the log-likelihood-ratio (G), the Freeman-Tukey/Hellinger distance, and other members of the Cressie-Read power-divergence family. The present paper indicates that the Euclidean/Frobenius/Hilbert-Schmidt distance can be more powerful for gauging the consistency of observed data with the assumption of homogeneity, too.