Prediction of outpatient rehabilitation patient preferences and optimization of graded diagnosis and treatment based on XGBoost machine learning algorithm

被引:0
|
作者
Fan, Xuehui [1 ]
Ye, Ruixue [1 ]
Gao, Yan [1 ]
Xue, Kaiwen [1 ]
Zhang, Zeyu [1 ]
Xu, Jing [1 ]
Zhao, Jingpu [1 ]
Feng, Jun [2 ]
Wang, Yulong [1 ]
机构
[1] Shenzhen Univ, Affiliated Hosp 1, Peoples Hosp Shenzhen 2, Dept Rehabil Med, Shenzhen, Guangdong, Peoples R China
[2] Linping Hosp Integrated Tradit Chinese & Western M, Hangzhou, Zhejiang, Peoples R China
来源
关键词
XGBoost; machine learning algorithm; rehabilitation patient; graded diagnosis and treatment; treatment preferences; HEALTH-CARE; CLASSIFICATION; CHALLENGES; MORTALITY; COVERAGE; CHINA;
D O I
10.3389/frai.2024.1473837
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Background The Department of Rehabilitation Medicine is key to improving patients' quality of life. Driven by chronic diseases and an aging population, there is a need to enhance the efficiency and resource allocation of outpatient facilities. This study aims to analyze the treatment preferences of outpatient rehabilitation patients by using data and a grading tool to establish predictive models. The goal is to improve patient visit efficiency and optimize resource allocation through these predictive models.Methods Data were collected from 38 Chinese institutions, including 4,244 patients visiting outpatient rehabilitation clinics. Data processing was conducted using Python software. The pandas library was used for data cleaning and preprocessing, involving 68 categorical and 12 continuous variables. The steps included handling missing values, data normalization, and encoding conversion. The data were divided into 80% training and 20% test sets using the Scikit-learn library to ensure model independence and prevent overfitting. Performance comparisons among XGBoost, random forest, and logistic regression were conducted using metrics, including accuracy and receiver operating characteristic (ROC) curves. The imbalanced learning library's SMOTE technique was used to address the sample imbalance during model training. The model was optimized using a confusion matrix and feature importance analysis, and partial dependence plots (PDP) were used to analyze the key influencing factors.Results XGBoost achieved the highest overall accuracy of 80.21% with high precision and recall in Category 1. random forest showed a similar overall accuracy. Logistic Regression had a significantly lower accuracy, indicating difficulties with nonlinear data. The key influencing factors identified include distance to medical institutions, arrival time, length of hospital stay, and specific diseases, such as cardiovascular, pulmonary, oncological, and orthopedic conditions. The tiered diagnosis and treatment tool effectively helped doctors assess patients' conditions and recommend suitable medical institutions based on rehabilitation grading.Conclusion This study confirmed that ensemble learning methods, particularly XGBoost, outperform single models in classification tasks involving complex datasets. Addressing class imbalance and enhancing feature engineering can further improve model performance. Understanding patient preferences and the factors influencing medical institution selection can guide healthcare policies to optimize resource allocation, improve service quality, and enhance patient satisfaction. Tiered diagnosis and treatment tools play a crucial role in helping doctors evaluate patient conditions and make informed recommendations for appropriate medical care.
引用
收藏
页数:18
相关论文
共 50 条
  • [21] Sleep prediction algorithm based on machine learning technology
    Park, K.
    Lee, S.
    Lee, S.
    Wang, S.
    Kim, S.
    Lee, E.
    JOURNAL OF SLEEP RESEARCH, 2018, 27
  • [22] Early Prediction of Sepsis Based on Machine Learning Algorithm
    Zhao, Xin
    Shen, Wenqian
    Wang, Guanjun
    COMPUTATIONAL INTELLIGENCE AND NEUROSCIENCE, 2021, 2021
  • [23] Prediction of self-consolidating concrete properties using XGBoost machine learning algorithm: Rheological properties
    Safhi, Amine El Mahdi
    Dabiri, Hamed
    Soliman, Ahmed
    Khayat, Kamal H.
    POWDER TECHNOLOGY, 2024, 438
  • [24] Machine learning and the optimization of prediction-based policies
    Battiston, Pietro
    Gamba, Simona
    Santoro, Alessandro
    TECHNOLOGICAL FORECASTING AND SOCIAL CHANGE, 2024, 199
  • [25] A Machine Learning Based Ensemble Forecasting Optimization Algorithm for Preseason Prediction of Atlantic Hurricane Activity
    Sun, Xia
    Xie, Lian
    Shah, Shahil Umeshkumar
    Shen, Xipeng
    ATMOSPHERE, 2021, 12 (04)
  • [26] Prediction & optimization of alkali-activated concrete based on the random forest machine learning algorithm
    Sun, Yubo
    Cheng, Hao
    Zhang, Shizhe
    Mohan, Manu K.
    Ye, Guang
    De Schutter, Geert
    CONSTRUCTION AND BUILDING MATERIALS, 2023, 385
  • [27] Structural deformation prediction model based on extreme learning machine algorithm and particle swarm optimization
    Jiang, Shouyan
    Zhao, Linxin
    Du, Chengbin
    STRUCTURAL HEALTH MONITORING-AN INTERNATIONAL JOURNAL, 2022, 21 (06): : 2786 - 2803
  • [28] Research on the Application of the Machine Learning Algorithm Based on Parameter Optimization in Network Security Situation Prediction
    Wang, Xiaoyan
    Wang, Jiangli
    International Journal of Network Security, 2023, 25 (02): : 245 - 251
  • [29] Prediction of patient choice tendency in medical decision-making based on machine learning algorithm
    Lyu, Yuwen
    Xu, Qian
    Yang, Zhenchao
    Liu, Junrong
    FRONTIERS IN PUBLIC HEALTH, 2023, 11
  • [30] A hybrid machine learning optimization algorithm for multivariable pore pressure prediction
    Song Deng
    HaoYu Pan
    HaiGe Wang
    ShouKun Xu
    XiaoPeng Yan
    ChaoWei Li
    MingGuo Peng
    HaoPing Peng
    Lin Shi
    Meng Cui
    Fei Zhao
    Petroleum Science, 2024, 21 (01) : 535 - 550