Evaluating Performance Metrics for Credit Card Fraud Classification

被引:7
|
作者
Leevy, Joffrey L. [1 ]
Khoshgoftaar, Taghi M. [1 ]
Hancock, John [1 ]
机构
[1] Florida Atlantic Univ, Boca Raton, FL 33431 USA
关键词
Extremely Randomized Trees; XGBoost; CatBoost; LightGBM; Random Forest; Class Imbalance; Undersampling; AUC; AUPRC; ALGORITHMS;
D O I
10.1109/ICTAI56018.2022.00202
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Practitioners and researchers of machine learning should have a deep understanding about the selection of the right performance metrics for classifier evaluation. Using a credit card fraud dataset, we demonstrate that the Area Under the Precision-Recall Curve (AUPRC) metric is a more reliable measurement, for the classification of highly imbalanced data, than the Area Under the Receiver Operating Characteristic Curve (AUC) metric. Furthermore, we establish that AUC is minimally impacted by the use of Random Undersampling (RUS). The classifiers used in this study are ensemble learners: LightGBM, CatBoost, Extremely Randomized Trees (ET), XGBoost, and Random Forest. Our results are governed by the fact that in a highly imbalanced dataset, the comparatively large number of true negative instances has an influence on AUC but not on AUPRC. Hence, AUPRC is able to accurately detect changes in the number of false positives because it ignores the true negatives.
引用
收藏
页码:1336 / 1341
页数:6
相关论文
共 50 条
  • [21] Performance Evaluation of Class Balancing Techniques for Credit Card Fraud Detection
    Sisodia, Dilip Singh
    Reddy, Nerella Keerthana
    Bhandari, Shivangi
    2017 IEEE INTERNATIONAL CONFERENCE ON POWER, CONTROL, SIGNALS AND INSTRUMENTATION ENGINEERING (ICPCSI), 2017, : 2747 - 2752
  • [22] Performance Evaluation of Machine Learning Algorithms for Credit Card Fraud Detection
    Mittal, Sangeeta
    Tyagi, Shivani
    2019 9TH INTERNATIONAL CONFERENCE ON CLOUD COMPUTING, DATA SCIENCE & ENGINEERING (CONFLUENCE 2019), 2019, : 320 - 324
  • [23] The Effects of Sampling Technique and Sample Size on Classification Performance of Ensemble Tree Classifiers in Detecting Credit Card Fraud
    Chogugudza, Mcdonald
    Ngwenya, Sikhumbuzo
    Sibanda, Khulumani
    2024 IST-AFRICA CONFERENCE, 2024,
  • [24] USING FRAUD TREES TO ANALYZE INTERNET CREDIT CARD FRAUD
    Blackwell, Clive
    ADVANCES IN DIGITAL FORENSICS X, 2014, 433 : 17 - 29
  • [25] Ensemble Learning for Credit Card Fraud Detection
    Sohony, Ishan
    Pratap, Rameshwar
    Nambiar, Ullas
    PROCEEDINGS OF THE ACM INDIA JOINT INTERNATIONAL CONFERENCE ON DATA SCIENCE AND MANAGEMENT OF DATA (CODS-COMAD'18), 2018, : 289 - 294
  • [26] Credit Card Fraud Detection with Attention Mechanism
    Han Lk
    Anh, L. T.
    PROCEEDINGS OF THE 2024 9TH INTERNATIONAL CONFERENCE ON INTELLIGENT INFORMATION TECHNOLOGY, ICIIT 2024, 2024, : 430 - 433
  • [27] Hybrid approaches for detecting credit card fraud
    Kultur, Yigit
    Caglayan, Mehmet Ufuk
    EXPERT SYSTEMS, 2017, 34 (02)
  • [28] Credit card fraud: Minerva's experience
    Dorronsoro, JR
    NEURAL NETWORKS: BEST PRACTICE IN EUROPE, 1997, 8 : 55 - 63
  • [29] Random Forest for Credit Card Fraud Detection
    Xuan, Shiyang
    Liu, Guanjun
    Li, Zhenchuan
    Zheng, Lutao
    Wang, Shuo
    Jiang, Changjun
    2018 IEEE 15TH INTERNATIONAL CONFERENCE ON NETWORKING, SENSING AND CONTROL (ICNSC), 2018,
  • [30] Using generative adversarial networks for improving classification effectiveness in credit card fraud detection
    Fiore, Ugo
    De Santis, Alfredo
    Perla, Francesca
    Zanetti, Paolo
    Palmieri, Francesco
    INFORMATION SCIENCES, 2019, 479 : 448 - 455