Evolution of Breast Cancer Recurrence Risk Prediction: A Systematic Review of Statistical and Machine Learning-Based Models

被引:2
|
作者
El Haji, Hasna [1 ,2 ,3 ]
Souadka, Amine [4 ]
Patel, Bhavik N. [1 ,2 ]
Sbihi, Nada
Ramasamy, Gokul [1 ,2 ]
Patel, Bhavika K. [1 ]
Ghogho, Mounir [3 ,5 ]
Banerjee, Imon [1 ,2 ]
机构
[1] Mayo Clin, Dept Radiol, 6161 E Mayo Blvd, Phoenix, AZ 85054 USA
[2] Arizona State Univ, Sch Comp & Augmented Intelligence, Tempe, AZ USA
[3] Int Univ Rabat, TICLab, Rabat, Morocco
[4] Mohammed V Univ Rabat, Natl Inst Oncol, Dept Surg Oncol, Rabat, Morocco
[5] Univ Leeds, Fac Engn, Leeds, W Yorkshire, England
来源
关键词
DISTANT RECURRENCE; ASSAY; EXPRESSION; NOMOGRAM; FEATURES; SURGERY;
D O I
10.1200/CCI.23.00049
中图分类号
R73 [肿瘤学];
学科分类号
100214 ;
摘要
PURPOSE Selection of appropriate adjuvant therapy to ultimately reduce the risk of breast cancer (BC) recurrence is a challenge for medical oncologists. Several automated risk prediction models have been developed using retrospective clinical data and have evolved significantly over the years in terms of predictors of recurrence, data usage, and predictive techniques (statistical/machine learning [ML]). METHODS Following PRISMA guidelines, we performed a systematic literature review of the aforementioned statistical and ML models published between January 2008 and December 2022 through searching five digital databases-PubMed, ScienceDirect, Scopus, Cochrane, and Web of Science. The comprehensive search yielded a total of 163 papers and after a screening process focusing on papers that dealt exclusively with statistical/ML methods, only 23 papers were deemed appropriate for further analysis. We benchmarked the studies on the basis of development, evaluation metrics, and validation strategy with an added emphasis on racial diversity of patients included in the studies. RESULTS In total, 30.4% of the included studies use statistical techniques, while 69.6% are ML-based. Among these, traditional ML models (support vector machines, decision tree, logistic regression, and naive Bayes) are the most frequently used (26.1%) along with deep learning (26.1%). Deep learning and ensemble learning provide the most accurate predictions (AUC = 0.94 each). CONCLUSION ML-based prediction models exhibit outstanding performance, yet their practical applicability might be hindered by limited interpretability and reduced generalization. Moreover, predictive models for BC recurrence often focus on limited variables related to tumor, treatment, molecular, and clinical features. Imbalanced classes and the lack of open-source data sets impede model development and validation. Furthermore, existing models predominantly overlook African and Middle Eastern populations, as they are trained and validated mainly on Caucasian and Asian patients.
引用
收藏
页数:14
相关论文
共 50 条
  • [31] Enhancing fairness in breast cancer recurrence prediction through temporal machine learning models
    Sundus, Katrina I.
    Hammo, Bassam H.
    Al-Zoubi, Mohammad B.
    Neural Computing and Applications, 2024, 36 (36) : 22697 - 22718
  • [32] Risk prediction models of breast cancer: a systematic review of model performances
    Anothaisintawee, Thunyarat
    Teerawattananon, Yot
    Wiratkapun, Chollathip
    Kasamesup, Vijj
    Thakkinstian, Ammarin
    BREAST CANCER RESEARCH AND TREATMENT, 2012, 133 (01) : 1 - 10
  • [33] Risk prediction models of breast cancer: a systematic review of model performances
    Thunyarat Anothaisintawee
    Yot Teerawattananon
    Chollathip Wiratkapun
    Vijj Kasamesup
    Ammarin Thakkinstian
    Breast Cancer Research and Treatment, 2012, 133 : 1 - 10
  • [34] Pre-existing and machine learning-based models for cardiovascular risk prediction
    Sang-Yeong Cho
    Sun-Hwa Kim
    Si-Hyuck Kang
    Kyong Joon Lee
    Dongjun Choi
    Seungjin Kang
    Sang Jun Park
    Tackeun Kim
    Chang-Hwan Yoon
    Tae-Jin Youn
    In-Ho Chae
    Scientific Reports, 11
  • [35] Pre-existing and machine learning-based models for cardiovascular risk prediction
    Cho, Sang-Yeong
    Kim, Sun-Hwa
    Kang, Si-Hyuck
    Lee, Kyong Joon
    Choi, Dongjun
    Kang, Seungjin
    Park, Sang Jun
    Kim, Tackeun
    Yoon, Chang-Hwan
    Youn, Tae-Jin
    Chae, In-Ho
    SCIENTIFIC REPORTS, 2021, 11 (01) : 8886
  • [36] Diagnostic Performance of Machine Learning-based Models in Neonatal Sepsis: A Systematic Review
    Kainth, Deepika
    Prakash, Satya
    Sankar, M. Jeeva
    PEDIATRIC INFECTIOUS DISEASE JOURNAL, 2024, 43 (09) : 889 - 901
  • [37] Machine Learning-Based Predictive Models for Patients with Venous Thromboembolism: A Systematic Review
    Danilatou, Vasiliki
    Dimopoulos, Dimitrios
    Kostoulas, Theodoros
    Douketis, James
    THROMBOSIS AND HAEMOSTASIS, 2024,
  • [38] Systematic review finds "spin"practices and poor reporting standards in studies on machine learning-based prediction models
    Navarro, Constanza L. Andaur
    Damen, Johanna A. A.
    Takada, Toshihiko
    Nijman, Steven W. J.
    Dhiman, Paula
    Ma, Jie
    Collins, Gary S.
    Bajpai, Ram
    Riley, Richard D.
    Moons, Karel G. M.
    Hooft, Lotty
    JOURNAL OF CLINICAL EPIDEMIOLOGY, 2023, 158 : 99 - 110
  • [39] Machine learning-based 30-day readmission prediction models for patients with heart failure: a systematic review
    Yu, Min-Young
    Son, Youn-Jung
    EUROPEAN JOURNAL OF CARDIOVASCULAR NURSING, 2024,
  • [40] Interpretability of machine learning-based prediction models in healthcare
    Stiglic, Gregor
    Kocbek, Primoz
    Fijacko, Nino
    Zitnik, Marinka
    Verbert, Katrien
    Cilar, Leona
    WILEY INTERDISCIPLINARY REVIEWS-DATA MINING AND KNOWLEDGE DISCOVERY, 2020, 10 (05)