Publiora

Menghubungkan ke Publiora...

Publiora

The effects of data imbalance on fraud detection model accuracy

Ruslan, Rusma AniezaArbaiy, NureizeLin, Pei-Chun
IAES International Journal of Artificial Intelligence (IJ-AI) (Sinta 1)Vol. 0 No. 01 April 2026
DOI10.11591/ijai.v15.i2.pp1402-1408

Abstrak

Machine learning (ML) model performance is often assessed by accuracy, but the quality and balance of data also play crucial roles. Imbalanced datasets, where the minority class has fewer samples than the majority class, can lead to biased predictions favoring the majority class. This study addresses the issue of class imbalance through resampling techniques, including random undersampling (RUS) and random oversampling (ROS), specifically applied to a fraud detection dataset. We classify the resampled datasets using random forest (RF) and gradient boosting (GB) models. Our findings indicate that the RF model, when combined with ROS, achieves an accuracy of 97.4%, surpassing the 96.1% accuracy of the GB model with RUS. This approach demonstrates the importance of addressing class imbalance to improve prediction accuracy in ML.

Kata Kunci

Data augmentationImbalanced datasetMachine learningResamplingSMOTE

Cari jurnal yang tepat untuk naskah Anda

MatchMind AI mencocokkan abstrak naskah Anda dengan ribuan jurnal terakreditasi dan menampilkan rekomendasi terbaik beserta alasannya.

Coba MatchMind

Lihat profil lengkap jurnal ini

Waktu review, biaya APC, statistik sitasi, indeksasi Scopus, dan banyak lagi.

Buka IAES International Journal of Artificial Intelligence (IJ-AI)

Artikel ini juga tersedia di situs resmi jurnal.

The effects of data imbalance on fraud detection model accuracy | IAES International Journal of Artificial Intelligence (IJ-AI) | Publiora