A stacked ensemble approach with resampling techniques for highly effective fraud detection in imbalanced datasets
Keywords:
Imbalanced dataset, Ensemble Approach, Fraud detection, Stacking algorithm, Synthetic Minority Oversampling Technique (SMOTE)Abstract
In several earlier studies, machine learning (ML) has been widely explored for fraud detection. However, fraud detection is still a challenging problem. This is due to the imbalanced nature of fraud data, which leads to underperformance by most models in detecting a few fraud cases. Undetected fraud cases also account for the loss of several millions of dollars annually. Thus, we propose an ensemble approach that stacks five classifiers - Support Vector Machine, Decision Trees, Random Forests, Gaussian Na¨?ve Bayes, and k-Nearest Neighbour, and uses the Logistic Regression meta-classifier to make predictions based on a stacking algorithm and novel pipeline. The effectiveness of the proposed model is examined on three datasets. The first two datasets were trained and tested initially without resampling and then compared with the results obtained using the Synthetic Minority Oversampling Technique (SMOTE) and RandomUnderSampler techniques. Only a balanced resampled dataset was trained on the third dataset that clearly showed an imbalance. From the results obtained, it is observed that the proposed model is highly competitive, with extant models producing ROC AUC of 99% and scoring above 98% in all other metrics. The approach is recommended for detecting fraud cases in similar case studies.
Published
How to Cite
Issue
Section
Copyright (c) 2024 Idongesit E. Eteng, Udeze L. Chinedu, Ayei E. Ibor

This work is licensed under a Creative Commons Attribution 4.0 International License.
How to Cite
Similar Articles
- L. G. Salaudeen, D. GABI, M. Garba, H. U. Suru, Deep convolutional neural network based synthetic minority over sampling technique: a forfending model for fraudulent credit card transactions in financial institution , Journal of the Nigerian Society of Physical Sciences: Volume 6, Issue 2, May 2024
- Chinedu L. Udeze, Idongesit E. Eteng, Ayei E. Ibor, Application of Machine Learning and Resampling Techniques to Credit Card Fraud Detection , Journal of the Nigerian Society of Physical Sciences: Volume 4, Issue 3, August 2022
- Raphael Ozighor Enihe, Rajesh Prasad, Francisca Nonyelum Ogwueleka, Fatimah Binta Abdullahi, The effect of imbalance data mitigation techniques on cardiovascular disease prediction , Journal of the Nigerian Society of Physical Sciences: Volume 7, Issue 2, May 2025
- Osowomuabe Njama-Abang, Denis U. Ashishie, Paul T. Bukie, Addressing class imbalance in lassa fever epidemic data, using machine learning: a case study with SMOTE and random forest , Journal of the Nigerian Society of Physical Sciences: Volume 7, Issue 3, August 2025
- Sherifdeen O. Bolarinwa, Eli Danladi, Andrew Ichoja, Muhammad Y. Onimisia, Christopher U. Achem, Synergistic Study of Reduced Graphene Oxide as Interfacial Buffer Layer in HTL-free Perovskite Solar Cells with Carbon Electrode , Journal of the Nigerian Society of Physical Sciences: Volume 4, Issue 3, August 2022
- Philemon Uten Emmoh, Christopher Ifeanyi Eke, Timothy Moses, A feature selection and scoring scheme for dimensionality reduction in a machine learning task , Journal of the Nigerian Society of Physical Sciences: Volume 7, Issue 1, February 2025
- Emmanuel Gbenga Dada, Aishatu Ibrahim Birma, Abdulkarim Abbas Gora, Ensemble machine learning algorithm for cost-effective and timely detection of diabetes in Maiduguri, Borno State , Journal of the Nigerian Society of Physical Sciences: Volume 6, Issue 4, November 2024
- Gerard Shu Fuhnwi, Janet O. Agbaje, Kayode Oshinubi, Olumuyiwa James Peter, An Empirical Study on Anomaly Detection Using Density-based and Representative-based Clustering Algorithms , Journal of the Nigerian Society of Physical Sciences: Volume 5, Issue 2, May 2023
- Akila Dabara Kayit, Mohd Tahir Ismail, Novel way to predict stock movements using multiple models and comprehensive analysis: leveraging voting meta-ensemble techniques , Journal of the Nigerian Society of Physical Sciences: Volume 6, Issue 3, August 2024
- Christian N. Nwaeme, Adewale F. Lukman, Robust hybrid algorithms for regularization and variable selection in QSAR studies , Journal of the Nigerian Society of Physical Sciences: Volume 5, Issue 4, November 2023
You may also start an advanced similarity search for this article.
Most read articles by the same author(s)
- Chinedu L. Udeze, Idongesit E. Eteng, Ayei E. Ibor, Application of Machine Learning and Resampling Techniques to Credit Card Fraud Detection , Journal of the Nigerian Society of Physical Sciences: Volume 4, Issue 3, August 2022

