The Impact of Feature Selection on Different Machine Learning Models for Breast Cancer Classification

被引:0
|
作者
Algherairy, Atheer [1 ,2 ]
Almattar, Wadha [1 ,2 ]
Bakri, Eman [2 ]
Albelali, Salma [2 ,3 ]
机构
[1] Imam Abdulrahman Bin Faisal Univ, Dept Comp Sci, Dammam, Saudi Arabia
[2] KFUPM, Informat & Comp Sci Dept, Dhahran, Saudi Arabia
[3] Imam Abdulrahman Bin Faisal Univ, Dept Comp Sci, Jubail Ind City, Saudi Arabia
关键词
Machine Learning; Feature Selection; Breast Cancer;
D O I
10.1109/CDMA54072.2022.00020
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Breast cancer appears to be a common type of cancer suffered by women globally, with considered high death rates. The survival rate of breast cancer patients decreases considerably for patients diagnosed at an advanced stage compared to those diagnosed at an early stage. The objective of this study is to investigate breast cancer classification and diagnosis task using the data from WBCD dataset. In our methodology, first, the breast cancer data was scaled. Then, four features selection methods were used to analyze the features. Pearson's Correlation method, Forward Selection method, Mutual Information and Univariate ROC-AUC were the used feature selectors. Next, different Machine Leaning models were applied including Support Vector Machine, Logistic Regression and XGBoost. Finally, the three models were cross-validated by 5-fold method. The ML models with different classifiers were evaluated based on several performance measures including accuracy, precision, recall, and F1-score. results show that Logistic Regression (LR) model with Forward Selection appeared to be the most successful classifier. The obtained classification accuracy, precision, and F1-score were 0.982, 0.983, 0.986; respectively. However, the highest recall score was 0.992 achieved by SVM model with Correlation feature selection. The developed model could potentially help the medical experts for the early diagnosis of breast cancer to decrease potential risk.
引用
收藏
页码:91 / 96
页数:6
相关论文
共 50 条
  • [1] Feature selection and classification in breast cancer prediction using IoT and machine learning
    Gopal, V. Nanda
    Al-Turjman, Fadi
    Kumar, R.
    Anand, L.
    Rajesh, M.
    [J]. MEASUREMENT, 2021, 178
  • [2] Feature selection and machine learning method for classification of lung cancer types
    Shin, Byungju
    Wang, Bohyun
    Lim, Joon S.
    [J]. Test Engineering and Management, 2019, 81 : 2307 - 2314
  • [3] Feature selection in a machine learning system for texture classification
    Baik, SW
    Bala, J
    [J]. ALGORITHMS FOR SYNTHETIC APERTURE RADAR IMAGERY V, 1998, 3370 : 261 - 268
  • [4] FEATURE SELECTION AND MACHINE LEARNING CLASSIFICATION FOR MALWARE DETECTION
    Khammas, Ban Mohammed
    Monemi, Alireza
    Bassi, Joseph Stephen
    Ismail, Ismahani
    Nor, Sulaiman Mohd
    Marsono, Muhammad Nadzir
    [J]. JURNAL TEKNOLOGI, 2015, 77 (01):
  • [5] A Comparative Study for Breast Cancer Prediction using Machine Learning and Feature Selection
    Dhanya, R.
    Paul, Irene Rose
    Akula, Sai Sindhu
    Sivakumar, Madhumathi
    Nair, Jyothisha J.
    [J]. PROCEEDINGS OF THE 2019 INTERNATIONAL CONFERENCE ON INTELLIGENT COMPUTING AND CONTROL SYSTEMS (ICCS), 2019, : 1049 - 1055
  • [6] Breast cancer classification along with feature prioritization using machine learning algorithms
    Abdullah-Al Nahid
    Md. Johir Raihan
    Abdullah Al-Mamun Bulbul
    [J]. Health and Technology, 2022, 12 : 1061 - 1069
  • [7] Comparison of Machine Learning Classifiers for Breast Cancer Diagnosis Based on Feature Selection
    Liu, Bo
    Li, Xingrui
    Li, Jianqiang
    Li, Yong
    Lang, Jianlei
    Gu, Rentao
    Wang, Fei
    [J]. 2018 IEEE INTERNATIONAL CONFERENCE ON SYSTEMS, MAN, AND CYBERNETICS (SMC), 2018, : 4385 - 4390
  • [8] Breast cancer classification along with feature prioritization using machine learning algorithms
    Abdullah-Al Nahid
    Raihan, Md Johir
    Bulbul, Abdullah Al-Mamun
    [J]. HEALTH AND TECHNOLOGY, 2022, 12 (06) : 1061 - 1069
  • [9] Feature selection with Fast Correlation-Based Filter for Breast cancer prediction and Classification using Machine Learning Algorithms
    Khourdifi, Youness
    Bahaj, Mohamed
    [J]. 2018 INTERNATIONAL SYMPOSIUM ON ADVANCED ELECTRICAL AND COMMUNICATION TECHNOLOGIES (ISAECT), 2018,
  • [10] Filter-Based Feature Selection and Machine-Learning Classification of Cancer Data
    Farsi, Mohammed
    [J]. INTELLIGENT AUTOMATION AND SOFT COMPUTING, 2021, 28 (01): : 83 - 92