vix.ing · top · new · best · stats · spec

Evaluating the relative contribution of data sources in a Bayesian analysis with the application of estimating the size of hard to reach populations

2020/09/04 by Jacob Parsons, Parsons, Jacob, Xiaoyue Niu +3
Computer Science · Mathematics · #Applications (stat.AP) #Bayesian Methods and Mixture Models #Census and Population Estimation #FOS: Computer and information sciences #Statistical Methods and Bayesian Inference

paper · pdf · doi:10.48550/arxiv.2009.02372

openalex publication_date 2020/09/04 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

When using multiple data sources in an analysis, it is important to understand the influence of each data source on the analysis and the consistency of the data sources with each other and the model. We suggest the use of a retrospective value of information framework in order to address such concerns. Value of information methods can be computationally difficult. We illustrate the use of computational methods that allow these methods to be applied even in relatively complicated settings. In illustrating the proposed methods, we focus on an application in estimating the size of hard to reach populations. Specifically, we consider estimating the number of injection drug users in Ukraine by combining all available data sources spanning over half a decade and numerous sub-national areas in the Ukraine. This application is of interest to public health researchers as this hard to reach population that plays a large role in the spread of HIV. We apply a Bayesian hierarchical model and evaluate the contribution of each data source in terms of absolute influence, expected influence, and level of surprise. Finally we apply value of information methods to inform suggestions on future data collection.

Citations

Related