Evaluating Performance Metrics for Credit Card Fraud Classification

被引：7

作者：

Leevy, Joffrey L. ^{[1
]}

Khoshgoftaar, Taghi M. ^{[1
]}

Hancock, John ^{[1
]}

机构：

[1] Florida Atlantic Univ, Boca Raton, FL 33431 USA

来源：

2022 IEEE 34TH INTERNATIONAL CONFERENCE ON TOOLS WITH ARTIFICIAL INTELLIGENCE, ICTAI | 2022年

关键词：

Extremely Randomized Trees; XGBoost; CatBoost; LightGBM; Random Forest; Class Imbalance; Undersampling; AUC; AUPRC; ALGORITHMS;

D O I：

10.1109/ICTAI56018.2022.00202

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Practitioners and researchers of machine learning should have a deep understanding about the selection of the right performance metrics for classifier evaluation. Using a credit card fraud dataset, we demonstrate that the Area Under the Precision-Recall Curve (AUPRC) metric is a more reliable measurement, for the classification of highly imbalanced data, than the Area Under the Receiver Operating Characteristic Curve (AUC) metric. Furthermore, we establish that AUC is minimally impacted by the use of Random Undersampling (RUS). The classifiers used in this study are ensemble learners: LightGBM, CatBoost, Extremely Randomized Trees (ET), XGBoost, and Random Forest. Our results are governed by the fact that in a highly imbalanced dataset, the comparatively large number of true negative instances has an influence on AUC but not on AUPRC. Hence, AUPRC is able to accurately detect changes in the number of false positives because it ignores the true negatives.

引用

页码：1336 / 1341

页数：6

共 50 条

[21] Performance Evaluation of Class Balancing Techniques for Credit Card Fraud Detection
Sisodia, Dilip Singh
Reddy, Nerella Keerthana
Bhandari, Shivangi
2017 IEEE INTERNATIONAL CONFERENCE ON POWER, CONTROL, SIGNALS AND INSTRUMENTATION ENGINEERING (ICPCSI), 2017, : 2747 - 2752
[22] Performance Evaluation of Machine Learning Algorithms for Credit Card Fraud Detection
Mittal, Sangeeta
Tyagi, Shivani
2019 9TH INTERNATIONAL CONFERENCE ON CLOUD COMPUTING, DATA SCIENCE & ENGINEERING (CONFLUENCE 2019), 2019, : 320 - 324
[23] The Effects of Sampling Technique and Sample Size on Classification Performance of Ensemble Tree Classifiers in Detecting Credit Card Fraud
Chogugudza, Mcdonald
Ngwenya, Sikhumbuzo
Sibanda, Khulumani
2024 IST-AFRICA CONFERENCE, 2024,
[24] USING FRAUD TREES TO ANALYZE INTERNET CREDIT CARD FRAUD
Blackwell, Clive
ADVANCES IN DIGITAL FORENSICS X, 2014, 433 : 17 - 29
[25] Ensemble Learning for Credit Card Fraud Detection
Sohony, Ishan
Pratap, Rameshwar
Nambiar, Ullas
PROCEEDINGS OF THE ACM INDIA JOINT INTERNATIONAL CONFERENCE ON DATA SCIENCE AND MANAGEMENT OF DATA (CODS-COMAD'18), 2018, : 289 - 294
[26] Credit Card Fraud Detection with Attention Mechanism
Han Lk
Anh, L. T.
PROCEEDINGS OF THE 2024 9TH INTERNATIONAL CONFERENCE ON INTELLIGENT INFORMATION TECHNOLOGY, ICIIT 2024, 2024, : 430 - 433
[27] Hybrid approaches for detecting credit card fraud
Kultur, Yigit
Caglayan, Mehmet Ufuk
EXPERT SYSTEMS, 2017, 34 (02)
[28] Credit card fraud: Minerva's experience
Dorronsoro, JR
NEURAL NETWORKS: BEST PRACTICE IN EUROPE, 1997, 8 : 55 - 63
[29] Random Forest for Credit Card Fraud Detection
Xuan, Shiyang
Liu, Guanjun
Li, Zhenchuan
Zheng, Lutao
Wang, Shuo
Jiang, Changjun
2018 IEEE 15TH INTERNATIONAL CONFERENCE ON NETWORKING, SENSING AND CONTROL (ICNSC), 2018,
[30] Using generative adversarial networks for improving classification effectiveness in credit card fraud detection
Fiore, Ugo
De Santis, Alfredo
Perla, Francesca
Zanetti, Paolo
Palmieri, Francesco
INFORMATION SCIENCES, 2019, 479 : 448 - 455

← 1 2 3 4 5 →