Estimation and correction of bias in network simulations based on respondent-driven sampling data

被引:1
|
作者
Zhu, Lin [1 ,2 ]
Menzies, Nicolas A. [1 ]
Wang, Jianing [3 ]
Linas, Benjamin P. [3 ,4 ]
Goodreau, Steven M. [5 ]
Salomon, Joshua A. [2 ]
机构
[1] Harvard TH Chan Sch Publ Hlth, Dept Global Hlth & Populat, Boston, MA USA
[2] Stanford Univ, Dept Med, Sch Med, Stanford, CA 94305 USA
[3] Boston Med Ctr, Dept Med, Infect Dis Sect, Boston, MA USA
[4] Boston Univ, Sch Publ Hlth, Dept Epidemiol, Boston, MA USA
[5] Stanford Univ, Univ Washington, Ctr Studies Demog & Ecol, Dept Epidemiol,Dept Anthropol, Seattle, WA USA
关键词
HEPATITIS-C TRANSMISSION; INJECTION-DRUG USERS; SOCIAL NETWORK; PREEXPOSURE PROPHYLAXIS; RISK; RECRUITMENT; INFERENCE; SNOWBALL; CENTERS; MODELS;
D O I
10.1038/s41598-020-63269-0
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
Respondent-driven sampling (RDS) is widely used for collecting data on hard-to-reach populations, including information about the structure of the networks connecting the individuals. Characterizing network features can be important for designing and evaluating health programs, particularly those that involve infectious disease transmission. While the validity of population proportions estimated from RDS-based datasets has been well studied, little is known about potential biases in inference about network structure from RDS. We developed a mathematical and statistical platform to simulate network structures with exponential random graph models, and to mimic the data generation mechanisms produced by RDS. We used this framework to characterize biases in three important network statistics - density/mean degree, homophily, and transitivity. Generalized linear models were used to predict the network statistics of the original network from the network statistics of the sample network and observable sample design features. We found that RDS may introduce significant biases in the estimation of density/mean degree and transitivity, and may exaggerate homophily when preferential recruitment occurs. Adjustments to network-generating statistics derived from the prediction models could substantially improve validity of simulated networks in terms of density, and could reduce bias in replicating mean degree, homophily, and transitivity from the original network.
引用
收藏
页数:11
相关论文
共 50 条
  • [1] ESTIMATION AND CORRECTION OF BIAS IN NETWORK SIMULATIONS BASED ON RESPONDENT-DRIVEN SAMPLING DATA
    Zhu, Lin
    Menzies, Nicolas A.
    Wang, Jianing
    Linas, Benjamin P.
    Goodreau, Steven M.
    Salomon, Joshua A.
    MEDICAL DECISION MAKING, 2020, 40 (01) : E297 - E299
  • [2] Estimation and correction of bias in network simulations based on respondent-driven sampling data
    Lin Zhu
    Nicolas A. Menzies
    Jianing Wang
    Benjamin P. Linas
    Steven M. Goodreau
    Joshua A. Salomon
    Scientific Reports, 10
  • [3] Asymptotic seed bias in respondent-driven sampling
    Yan, Yuling
    Hanlon, Bret
    Roch, Sebastien
    Rohe, Karl
    ELECTRONIC JOURNAL OF STATISTICS, 2020, 14 (01): : 1577 - 1610
  • [4] Bias decomposition and estimator performance in respondent-driven sampling
    Sirianni, Antonio D.
    Cameron, Christopher J.
    Shi, Yongren
    Heckathorn, Douglas D.
    SOCIAL NETWORKS, 2021, 64 : 109 - 121
  • [5] Respondent-driven sampling
    Schonlau, Matthias
    Liebau, Elisabeth
    STATA JOURNAL, 2012, 12 (01): : 72 - 93
  • [6] Sampling and estimation in hidden populations using respondent-driven sampling
    Salganik, MJ
    Heckathorn, DD
    SOCIOLOGICAL METHODOLOGY, 2004, VOL 34, 2004, 34 : 193 - 239
  • [7] Network Sampling: From Snowball and Multiplicity to Respondent-Driven Sampling
    Heckathorn, Douglas D.
    Cameron, Christopher J.
    ANNUAL REVIEW OF SOCIOLOGY, VOL 43, 2017, 43 : 101 - 119
  • [8] Hidden Population Size Estimation From Respondent-Driven Sampling: A Network Approach
    Crawford, Forrest W.
    Wu, Jiacheng
    Heimer, Robert
    JOURNAL OF THE AMERICAN STATISTICAL ASSOCIATION, 2018, 113 (522) : 755 - 766
  • [9] Assessing respondent-driven sampling
    Goel, Sharad
    Salganik, Matthew J.
    PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2010, 107 (15) : 6743 - 6747
  • [10] Improved Inference for Respondent-Driven Sampling Data With Application to HIV Prevalence Estimation
    Gile, Krista J.
    JOURNAL OF THE AMERICAN STATISTICAL ASSOCIATION, 2011, 106 (493) : 135 - 146