A neural network-based framework for the reconstruction of incomplete data sets

被引:67
|
作者
Gheyas, Iffat A. [1 ]
Smith, Leslie S. [1 ]
机构
[1] Univ Stirling, Dept Comp Sci & Math, Stirling FK9 4LA, Scotland
关键词
Missing values; Imputation; Single imputation; Multiple imputation; Generalized regression neural networks; MULTIPLE IMPUTATION; GENERAL REGRESSION; GENETIC ALGORITHM; ENSEMBLE; CLASSIFIERS;
D O I
10.1016/j.neucom.2010.06.021
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
The treatment of incomplete data is an important step in the pre-processing of data. We propose a novel nonparametric algorithm Generalized regression neural network Ensemble for Multiple Imputation (GEMI). We also developed a single imputation (SI) version of this approach-GESI. We compare our algorithms with 25 popular missing data imputation algorithms on 98 real-world and synthetic datasets for various percentage of missing values. The effectiveness of the algorithms is evaluated in terms of (i) the accuracy of output classification: three classifiers (a generalized regression neural network, a multilayer perceptron and a logistic regression technique) are separately trained and tested on the dataset imputed with each imputation algorithm, (ii) interval analysis with missing observations and (iii) point estimation accuracy of the missing value imputation. GEMI outperformed GESI and all the conventional imputation algorithms in terms of all three criteria considered. (C) 2010 Elsevier B.V. All rights reserved.
引用
收藏
页码:3039 / 3065
页数:27
相关论文
共 50 条
  • [1] Neural Network-based Framework for Data Stream Mining
    Silva, Bruno
    Marques, Nuno
    [J]. PROCEEDINGS OF THE SIXTH STARTING AI RESEARCHERS' SYMPOSIUM (STAIRS 2012), 2012, 241 : 294 - +
  • [2] DEEP NEURAL NETWORK-BASED DATA RECONSTRUCTION FOR LANDSLIDE DETECTION
    Utomo, Darmawan
    Hu, Liang-Cheng
    Hsiung, Pao-Ann
    [J]. IGARSS 2020 - 2020 IEEE INTERNATIONAL GEOSCIENCE AND REMOTE SENSING SYMPOSIUM, 2020, : 3119 - 3122
  • [3] Incomplete points cloud data surface reconstruction based on neural network
    Wu Xue-mei
    Li Gui-xian
    Zhao Wei-min
    [J]. 2008 FOURTH INTERNATIONAL CONFERENCE ON INTELLIGENT INFORMATION HIDING AND MULTIMEDIA SIGNAL PROCESSING, PROCEEDINGS, 2008, : 913 - +
  • [4] Neural Network-Based Automatic Reconstruction of Missing Vessel Trajectory Data
    Liang, Maohan
    Liu, Ryan Wen
    Zhong, Qianru
    Liu, Jingxian
    Zhang, Jinfeng
    [J]. 2019 4TH IEEE INTERNATIONAL CONFERENCE ON BIG DATA ANALYTICS (ICBDA 2019), 2019, : 426 - 430
  • [5] Neural network-based processing and reconstruction of compromised biophotonic image data
    Fanous, Michael John
    Casteleiro Costa, Paloma
    Isil, Cagatay
    Huang, Luzhe
    Ozcan, Aydogan
    [J]. LIGHT-SCIENCE & APPLICATIONS, 2024, 13 (01)
  • [6] Neural network-based PET image reconstruction
    Kosugi, Y
    Sase, M
    Suganami, Y
    Uemoto, N
    Momose, T
    Nishikawa, J
    [J]. METHODS OF INFORMATION IN MEDICINE, 1997, 36 (4-5) : 329 - 331
  • [7] Deep neural network-based spatiotemporal heterogeneous data reconstruction for landslide detection
    Darmawan Utomo
    Liang-Cheng Hu
    Pao-Ann Hsiung
    [J]. International Journal of Data Science and Analytics, 2024, 17 : 93 - 109
  • [8] Deep neural network-based spatiotemporal heterogeneous data reconstruction for landslide detection
    Utomo, Darmawan
    Hu, Liang-Cheng
    Hsiung, Pao-Ann
    [J]. INTERNATIONAL JOURNAL OF DATA SCIENCE AND ANALYTICS, 2024, 17 (01) : 93 - 109
  • [9] Recurrent Neural Network-Based Approach for Sparse Geomagnetic Data Interpolation and Reconstruction
    Liu, Huan
    Liu, Zheng
    Dong, Haobin
    Ge, Jian
    Yuan, Zhiwen
    Zhu, Jun
    Zhang, Haiyang
    Zeng, Xuming
    [J]. IEEE ACCESS, 2019, 7 : 33173 - 33179
  • [10] A survey of network-based intrusion detection data sets
    Ring, Markus
    Wunderlich, Sarah
    Scheuring, Deniz
    Landes, Dieter
    Hotho, Andreas
    [J]. COMPUTERS & SECURITY, 2019, 86 : 147 - 167