An interpretable hybrid predictive model of COVID-19 cases using autoregressive model and LSTM

被引:6
|
作者
Zhang, Yangyi [1 ]
Tang, Sui [1 ]
Yu, Guo [2 ]
机构
[1] Univ Calif Santa Barbara, Dept Math, Santa Barbara, CA 93106 USA
[2] Univ Calif Santa Barbara, Dept Stat & Appl Probabil, Santa Barbara, CA 93106 USA
关键词
ARIMA; XGBOOST;
D O I
10.1038/s41598-023-33685-z
中图分类号
O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];
学科分类号
07 ; 0710 ; 09 ;
摘要
The Coronavirus Disease 2019 (COVID-19) has had a profound impact on global health and economy, making it crucial to build accurate and interpretable data-driven predictive models for COVID-19 cases to improve public policy making. The extremely large scale of the pandemic and the intrinsically changing transmission characteristics pose a great challenge for effectively predicting COVID-19 cases. To address this challenge, we propose a novel hybrid model in which the interpretability of the Autoregressive model (AR) and the predictive power of the long short-term memory neural networks (LSTM) join forces. The proposed hybrid model is formalized as a neural network with an architecture that connects two composing model blocks, of which the relative contribution is decided data-adaptively in the training procedure. We demonstrate the favorable performance of the hybrid model over its two single composing models as well as other popular predictive models through comprehensive numerical studies on two data sources under multiple evaluation metrics. Specifically, in county-level data of 8 California counties, our hybrid model achieves 4.173% MAPE, outperforming the composing AR (5.629%) and LSTM (4.934%) alone on average. In country-level datasets, our hybrid model outperforms the widely-used predictive models such as AR, LSTM, Support Vector Machines, Gradient Boosting, and Random Forest, in predicting the COVID-19 cases in Japan, Canada, Brazil, Argentina, Singapore, Italy, and the United Kingdom. In addition to the predictive performance, we illustrate the interpretability of our proposed hybrid model using the estimated AR component, which is a key feature that is not shared by most black-box predictive models for COVID-19 cases. Our study provides a new and promising direction for building effective and interpretable data-driven models for COVID-19 cases, which could have significant implications for public health policy making and control of the current COVID-19 and potential future pandemics.
引用
收藏
页数:12
相关论文
共 50 条
  • [41] Uncertain SEIAR model for COVID-19 cases in China
    Lifen Jia
    Wei Chen
    Fuzzy Optimization and Decision Making, 2021, 20 : 243 - 259
  • [42] Predictive analytics of COVID-19 cases and tourist arrivals in ASEAN based on covid-19 cases
    Velu, Shubashini Rathina
    Ravi, Vinayakumar
    Tabianan, Kayalvily
    HEALTH AND TECHNOLOGY, 2022, 12 (06) : 1237 - 1258
  • [43] Predictive analytics of COVID-19 cases and tourist arrivals in ASEAN based on covid-19 cases
    Shubashini Rathina Velu
    Vinayakumar Ravi
    Kayalvily Tabianan
    Health and Technology, 2022, 12 : 1237 - 1258
  • [44] Uncertain SEIAR model for COVID-19 cases in China
    Jia, Lifen
    Chen, Wei
    FUZZY OPTIMIZATION AND DECISION MAKING, 2021, 20 (02) : 243 - 259
  • [45] Modeling COVID-19 Using a Modified SVIR Compartmental Model and LSTM-Estimated Parameters
    Wyss, Alejandra
    Hidalgo, Arturo
    MATHEMATICS, 2023, 11 (06)
  • [46] Analysis and Prediction of COVID-19 by using Recurrent LSTM Neural Network Model in Machine Learning
    Dharani, N. P.
    Bojja, Polaiah
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2022, 13 (05) : 171 - 178
  • [47] COVID-19 Pandemic Forecasting Using CNN-LSTM: A Hybrid Approach
    Zain, Zuhaira M.
    Alturki, Nazik M.
    JOURNAL OF CONTROL SCIENCE AND ENGINEERING, 2021, 2021
  • [48] COVID-19 Prediction Classifier Model Using Hybrid Algorithms in Data Mining
    Nikooghadam, Morteza
    Ghazikhani, Adel
    Saeedi, Mohammad
    INTERNATIONAL JOURNAL OF PEDIATRICS-MASHHAD, 2021, 9 (01): : 12723 - 12737
  • [49] COVID-19 BEDS OCCUPANCY AND HOSPITAL COMPLAINTS: A PREDICTIVE MODEL
    Foglia, E.
    Ferrario, L. B.
    Bellavia, D.
    Schettini, F.
    Falletti, E.
    Gallese, C.
    Nobile, M. S.
    Riva, S. G.
    VALUE IN HEALTH, 2022, 25 (12) : S356 - S356
  • [50] Predictive model with analysis of the initial spread of COVID-19 in India
    Ghosh, Shinjini
    INTERNATIONAL JOURNAL OF MEDICAL INFORMATICS, 2020, 143