Prediction of outpatient rehabilitation patient preferences and optimization of graded diagnosis and treatment based on XGBoost machine learning algorithm

被引:0
|
作者
Fan, Xuehui [1 ]
Ye, Ruixue [1 ]
Gao, Yan [1 ]
Xue, Kaiwen [1 ]
Zhang, Zeyu [1 ]
Xu, Jing [1 ]
Zhao, Jingpu [1 ]
Feng, Jun [2 ]
Wang, Yulong [1 ]
机构
[1] Shenzhen Univ, Affiliated Hosp 1, Peoples Hosp Shenzhen 2, Dept Rehabil Med, Shenzhen, Guangdong, Peoples R China
[2] Linping Hosp Integrated Tradit Chinese & Western M, Hangzhou, Zhejiang, Peoples R China
来源
关键词
XGBoost; machine learning algorithm; rehabilitation patient; graded diagnosis and treatment; treatment preferences; HEALTH-CARE; CLASSIFICATION; CHALLENGES; MORTALITY; COVERAGE; CHINA;
D O I
10.3389/frai.2024.1473837
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Background The Department of Rehabilitation Medicine is key to improving patients' quality of life. Driven by chronic diseases and an aging population, there is a need to enhance the efficiency and resource allocation of outpatient facilities. This study aims to analyze the treatment preferences of outpatient rehabilitation patients by using data and a grading tool to establish predictive models. The goal is to improve patient visit efficiency and optimize resource allocation through these predictive models.Methods Data were collected from 38 Chinese institutions, including 4,244 patients visiting outpatient rehabilitation clinics. Data processing was conducted using Python software. The pandas library was used for data cleaning and preprocessing, involving 68 categorical and 12 continuous variables. The steps included handling missing values, data normalization, and encoding conversion. The data were divided into 80% training and 20% test sets using the Scikit-learn library to ensure model independence and prevent overfitting. Performance comparisons among XGBoost, random forest, and logistic regression were conducted using metrics, including accuracy and receiver operating characteristic (ROC) curves. The imbalanced learning library's SMOTE technique was used to address the sample imbalance during model training. The model was optimized using a confusion matrix and feature importance analysis, and partial dependence plots (PDP) were used to analyze the key influencing factors.Results XGBoost achieved the highest overall accuracy of 80.21% with high precision and recall in Category 1. random forest showed a similar overall accuracy. Logistic Regression had a significantly lower accuracy, indicating difficulties with nonlinear data. The key influencing factors identified include distance to medical institutions, arrival time, length of hospital stay, and specific diseases, such as cardiovascular, pulmonary, oncological, and orthopedic conditions. The tiered diagnosis and treatment tool effectively helped doctors assess patients' conditions and recommend suitable medical institutions based on rehabilitation grading.Conclusion This study confirmed that ensemble learning methods, particularly XGBoost, outperform single models in classification tasks involving complex datasets. Addressing class imbalance and enhancing feature engineering can further improve model performance. Understanding patient preferences and the factors influencing medical institution selection can guide healthcare policies to optimize resource allocation, improve service quality, and enhance patient satisfaction. Tiered diagnosis and treatment tools play a crucial role in helping doctors evaluate patient conditions and make informed recommendations for appropriate medical care.
引用
收藏
页数:18
相关论文
共 50 条
  • [41] MLMD-A Malware-Detecting Antivirus Tool Based on the XGBoost Machine Learning Algorithm
    Palsa, Jakub
    Adam, Norbert
    Hurtuk, Jan
    Chovancova, Eva
    Mados, Branislav
    Chovanec, Martin
    Kocan, Stanislav
    APPLIED SCIENCES-BASEL, 2022, 12 (13):
  • [42] Performance optimization of Bloch surface wave based devices using a XGBoost machine learning model
    Yi, Hongxian
    Goyal, Amit Kumar
    Massoud, Yehia
    OPTICS CONTINUUM, 2024, 3 (05): : 693 - 703
  • [43] Deformation Prediction of Concrete Dam Based on Improved Particle Swarm Optimization Algorithm and Extreme Learning Machine
    Li M.
    Wang J.
    Wang Y.
    Tianjin Daxue Xuebao (Ziran Kexue yu Gongcheng Jishu Ban)/Journal of Tianjin University Science and Technology, 2019, 52 (11): : 1136 - 1144
  • [44] Traffic Flow Prediction Method Based on Improved Optimization Algorithm Mixed Kernel Extreme Learning Machine
    Yu, Jianyou
    Feng, Zhigang
    Ju, Liang
    Pan, Xiu
    Ge, Jiqiang
    Chen, Yusen
    PROCEEDINGS OF INTERNATIONAL CONFERENCE ON ALGORITHMS, SOFTWARE ENGINEERING, AND NETWORK SECURITY, ASENS 2024, 2024, : 54 - 58
  • [45] Machine Learning Based Malaria Prediction Using KNN Algorithm
    Meenu, M.
    Subhasri, G.
    Mahima, R.
    BIOSCIENCE BIOTECHNOLOGY RESEARCH COMMUNICATIONS, 2020, 13 (02): : 38 - 42
  • [46] Prediction of Type 2 Diabetes Based on Machine Learning Algorithm
    Deberneh, Henock M.
    Kim, Intaek
    INTERNATIONAL JOURNAL OF ENVIRONMENTAL RESEARCH AND PUBLIC HEALTH, 2021, 18 (06)
  • [47] Machine learning in the septic surgical patient: a systematic review of postoperative sepsis diagnosis and prediction machine learning algorithms
    Boyer, Richard
    Armoundas, Antonis
    ANESTHESIA AND ANALGESIA, 2020, 130 : 244 - 246
  • [49] Machine learning-based classification of varicocoele grading: A promising approach for diagnosis and treatment optimization
    Kayra, Mehmet Vehbi
    Sahin, Ali
    Toksoz, Serdar
    Serindere, Mehmet
    Altintas, Emre
    Ozer, Halil
    Gul, Murat
    ANDROLOGY, 2024,
  • [50] Optimization of a Dynamic Fault Diagnosis Model Based on Machine Learning
    Zhang, Shigang
    Luo, Xu
    Yang, Yongmin
    Wang, Long
    Zhang, Xiaofei
    IEEE ACCESS, 2018, 6 : 65065 - 65077