Deeply Learned Generalized Linear Models with Missing Data

被引:0
|
作者
Lim, David K. [1 ]
Rashid, Naim U. [1 ]
Oliva, Junier B. [2 ]
Ibrahim, Joseph G. [1 ]
机构
[1] Univ North Carolina Chapel Hill, Dept Biostat, Chapel Hill, NC 27515 USA
[2] Univ North Carolina Chapel Hill, Dept Comp Sci, Chapel Hill, NC USA
关键词
Deeply learned glm; Missing data; MNAR; Supervised learning; IMPUTATION;
D O I
10.1080/10618600.2023.2276122
中图分类号
O21 [概率论与数理统计]; C8 [统计学];
学科分类号
020208 ; 070103 ; 0714 ;
摘要
Deep Learning (DL) methods have dramatically increased in popularity in recent years, with significant growth in their application to various supervised learning problems. However, the greater prevalence and complexity of missing data in such datasets present significant challenges for DL methods. Here, we provide a formal treatment of missing data in the context of deeply learned generalized linear models, a supervised DL architecture for regression and classification problems. We propose a new architecture, dlglm, that is one of the first to be able to flexibly account for both ignorable and non-ignorable patterns of missingness in input features and response at training time. We demonstrate through statistical simulation that our method outperforms existing approaches for supervised learning tasks in the presence of missing not at random (MNAR) missingness. We conclude with a case study of the Bank Marketing dataset from the UCI Machine Learning Repository, in which we predict whether clients subscribed to a product based on phone survey data. Supplementary materials for this article are available online.
引用
收藏
页码:638 / 650
页数:13
相关论文
共 50 条
  • [21] Generalized Linear Models for Insurance Data
    Thomas, Ian
    ANNALS OF ACTUARIAL SCIENCE, 2009, 4 (02) : 345 - 346
  • [22] Generalized Linear Models for Aggregated Data
    Bhowmik, Avradeep
    Ghosh, Joydeep
    Koyejo, Oluwasanmi
    ARTIFICIAL INTELLIGENCE AND STATISTICS, VOL 38, 2015, 38 : 93 - 101
  • [23] Model averaging for generalized linear models with missing at random covariates
    Cheng, Weili
    Li, Xiaorui
    Li, Xiaoxia
    Yan, Xiaodong
    STATISTICS, 2023, 57 (01) : 26 - 52
  • [24] Model selection of generalized partially linear models with missing covariates
    Fu, Ying-Zi
    Chen, Xue-Dong
    JOURNAL OF STATISTICAL PLANNING AND INFERENCE, 2012, 142 (01) : 126 - 138
  • [25] Bayesian analysis for generalized linear models with nonignorably missing covariates
    Huang, L
    Chen, MH
    Ibrahim, JG
    BIOMETRICS, 2005, 61 (03) : 767 - 780
  • [26] Generalized linear mixed models with informative dropouts and missing covariates
    Wu, Kunling
    Wu, Lang
    METRIKA, 2007, 66 (01) : 1 - 18
  • [27] Non-ignorable missing covariates in generalized linear models
    Lipsitz, SR
    Ibrahim, JG
    Chen, MH
    Peterson, H
    STATISTICS IN MEDICINE, 1999, 18 (17-18) : 2435 - 2448
  • [28] Robust methods for generalized linear models with nonignorable missing covariates
    Sinha, Sanjoy K.
    CANADIAN JOURNAL OF STATISTICS-REVUE CANADIENNE DE STATISTIQUE, 2008, 36 (02): : 277 - 299
  • [29] Bayesian methods for generalized linear models with covariates missing at random
    Ibrahim, JG
    Chen, MH
    Lipsitz, SR
    CANADIAN JOURNAL OF STATISTICS-REVUE CANADIENNE DE STATISTIQUE, 2002, 30 (01): : 55 - 78
  • [30] Generalized linear mixed models with informative dropouts and missing covariates
    Kunling Wu
    Lang Wu
    Metrika, 2007, 66 : 1 - 18