Multiple imputation with missing indicators as proxies for unmeasured variables: simulation study

被引:22
|
作者
Sperrin, Matthew [1 ]
Martin, Glen P. [1 ]
机构
[1] Univ Manchester, Fac Biol Med & Hlth, Vaughan House, Manchester M13 9PL, Lancs, England
关键词
Missing data; Missing indicator; Multiple imputation; Simulation study; REGRESSION; BIAS;
D O I
10.1186/s12874-020-01068-x
中图分类号
R19 [保健组织与事业(卫生事业管理)];
学科分类号
摘要
Background Within routinely collected health data, missing data for an individual might provide useful information in itself. This occurs, for example, in the case of electronic health records, where the presence or absence of data is informative. While the naive use of missing indicators to try to exploit such information can introduce bias, its use in conjunction with multiple imputation may unlock the potential value of missingness to reduce bias in causal effect estimation, particularly in missing not at random scenarios and where missingness might be associated with unmeasured confounders. Methods We conducted a simulation study to determine when the use of a missing indicator, combined with multiple imputation, would reduce bias for causal effect estimation, under a range of scenarios including unmeasured variables, missing not at random, and missing at random mechanisms. We use directed acyclic graphs and structural models to elucidate a variety of causal structures of interest. We handled missing data using complete case analysis, and multiple imputation with and without missing indicator terms. Results We find that multiple imputation combined with a missing indicator gives minimal bias for causal effect estimation in most scenarios. In particular the approach: 1) does not introduce bias in missing (completely) at random scenarios; 2) reduces bias in missing not at random scenarios where the missing mechanism depends on the missing variable itself; and 3) may reduce or increase bias when unmeasured confounding is present. Conclusion In the presence of missing data, careful use of missing indicators, combined with multiple imputation, can improve causal effect estimation when missingness is informative, and is not detrimental when missingness is at random.
引用
收藏
页数:11
相关论文
共 50 条
  • [31] Imputation of missing diagnosis date and implications for analyzes in registries: A simulation study
    Wilker, Elissa H.
    Rodrigues, Ema
    Cole, J. Alexander
    [J]. PHARMACOEPIDEMIOLOGY AND DRUG SAFETY, 2022, 31 : 553 - 554
  • [32] Missing Data in Marginal Structural Models A Plasmode Simulation Study Comparing Multiple Imputation and Inverse Probability Weighting
    Liu, Shao-Hsien
    Chrysanthopoulou, Stavroula A.
    Chang, Qiuzhi
    Hunnicutt, Jacob N.
    Lapane, Kate L.
    [J]. MEDICAL CARE, 2019, 57 (03) : 237 - 243
  • [33] Multiple imputation for survey data that are missing by design: A validation study.
    Yost, K
    Levine, R
    Gold, E
    [J]. AMERICAN JOURNAL OF EPIDEMIOLOGY, 2003, 157 (11) : S34 - S34
  • [34] Clustering with missing and left-censored data: A simulation study comparing multiple-imputation-based procedures
    Faucheux, Lilith
    Resche-Rigon, Matthieu
    Curis, Emmanuel
    Soumelis, Vassili
    Chevret, Sylvie
    [J]. BIOMETRICAL JOURNAL, 2021, 63 (02) : 372 - 393
  • [35] Multiple imputation for missing data: a brief introduction
    Baccini, Michela
    [J]. EPIDEMIOLOGIA & PREVENZIONE, 2008, 32 (03): : 162 - 163
  • [36] Multiple imputation of unordered categorical missing data: A comparison of the multivariate normal imputation and multiple imputation by chained equations
    Karangwa, Innocent
    Kotze, Danelle
    Blignaut, Renette
    [J]. BRAZILIAN JOURNAL OF PROBABILITY AND STATISTICS, 2016, 30 (04) : 521 - 539
  • [37] Outcome-sensitive multiple imputation: a simulation study
    Evangelos Kontopantelis
    Ian R. White
    Matthew Sperrin
    Iain Buchan
    [J]. BMC Medical Research Methodology, 17
  • [38] The use of multiple imputation for the analysis of missing data
    Sinharay, S
    Stern, HS
    Russell, D
    [J]. PSYCHOLOGICAL METHODS, 2001, 6 (04) : 317 - 329
  • [39] Multiple imputation for missing data - A cautionary tale
    Allison, PD
    [J]. SOCIOLOGICAL METHODS & RESEARCH, 2000, 28 (03) : 301 - 309
  • [40] Introduction to multiple imputation for dealing with missing data
    Lee, Katherine J.
    Simpson, Julie A.
    [J]. RESPIROLOGY, 2014, 19 (02) : 162 - 167