Explainable Ai-Based Fraud Detection in Fintech Applications
Authors
Department of Computer Science, McPherson University, Seriki Sotayo, Ogun State (Nigeria)
Department of Computer Science, Ajayi Crowther University, Oyo, Oyo State (Nigeria)
Department of Computer Science, Federal University of Agricultural Science, Abeokuta, Ogun State (Nigeria)
Department of Software Engineering, McPherson University, Seriki Sotayo, Ogun State (Nigeria)
Article Information
DOI: 10.51244/IJRSI.2026.1306000158
Subject Category: Explainable AI
Volume/Issue: 13/6 | Page No: 2087-2108
Publication Timeline
Submitted: 2026-06-06
Accepted: 2026-06-11
Published: 2026-06-29
Abstract
The rapid growth of digital financial services has significantly increased the volume of online transactions, making fraud detection a critical challenge for financial institutions. Traditional machine learning models often provide strong predictive performance but lack interpretability, limiting trust and practical adoption in financial decision-making. This study proposes an Explainable Artificial Intelligence (XAI)-based fraud detection framework for FinTech transactions using the Kaggle Credit Card Fraud Detection dataset containing 284,807 transactions, including 492 fraudulent cases. To address severe class imbalance, Synthetic Minority Oversampling Technique (SMOTE) was applied, increasing the dataset to 568,630 balanced instances. Data preprocessing involved feature scaling and train–test splitting prior to model training. Three machine learning algorithms—Logistic Regression, Random Forest, and Extreme Gradient Boosting (XGBoost) were developed and evaluated using accuracy, precision, recall, F1-score, and ROC-AUC metrics.
The experimental results demonstrate strong predictive performance across all models. Logistic Regression achieved 94.50% accuracy, 97.32% precision, 91.51% recall, 94.33% F1-score, and a ROC-AUC of 94.50%. Random Forest produced the highest overall performance with 99.99% accuracy, 99.98% precision, 100.00% recall, 99.99% F1-score, and 99.99% ROC-AUC. XGBoost also achieved excellent results with 99.97% accuracy, 99.94% precision, 100.00% recall, 99.97% F1-score, and 99.97% ROC-AUC. To improve model transparency, SHapley Additive exPlanations (SHAP) and Local Interpretable Model-Agnostic Explanations (LIME) were integrated with the XGBoost model to provide both global and local interpretability. SHAP analysis identified transaction amount and several transformed principal component features as the most influential predictors of fraudulent behavior, while LIME provided instance-level explanations for individual fraud predictions. Feature importance analysis from Random Forest and XGBoost further validated the consistency of the most influential variables.
The findings demonstrate that combining high-performing machine learning models with explainable AI techniques can significantly enhance fraud detection accuracy while maintaining transparency and interpretability. The proposed framework offers a reliable and practical approach for intelligent fraud prevention in financial technology systems and supports trustworthy decision-making in real-world financial environments.
Keywords
Explainable Artificial Intelligence (XAI), Fraud Detection, FinTech, Machine Learning, Random Forest, XGBoost, SHAP, LIME, Credit Card Fraud Detection
Downloads
References
1. A. Adadi and M. Berrada, “Peeking inside the black-box: A survey on explainable artificial intelligence (XAI),” IEEE Access, vol. 6, pp. 52138–52160, 2018. [Google Scholar] [Crossref]
2. A. Ahmad, M. Saad, and A. Alshamrani, “Credit card fraud detection using ensemble machine learning techniques,” IEEE Access, vol. 10, pp. 12345–12360, 2022. [Google Scholar] [Crossref]
3. A. Barredo Arrieta, N. Díaz-Rodríguez, J. Del Ser et al., “Explainable artificial intelligence (XAI): Concepts, taxonomies, opportunities and challenges toward responsible AI,” Information Fusion, vol. 58, pp. 82–115, 2020. [Google Scholar] [Crossref]
4. M. Buda, A. Maki, and M. A. Mazurowski, “A systematic study of the class imbalance problem in convolutional neural networks,” Neural Networks, vol. 106, pp. 249–259, 2020. [Google Scholar] [Crossref]
5. N. Bussmann, P. Giudici, D. Marinelli, and J. Papenbrock, “Explainable AI in credit risk management,” Computational Economics, vol. 57, pp. 203–216, 2021. [Google Scholar] [Crossref]
6. V. Chandola, A. Banerjee, and V. Kumar, “Anomaly detection: A survey,” ACM Computing Surveys, vol. 41, no. 3, pp. 1–58, 2009. [Google Scholar] [Crossref]
7. T. Chen and C. Guestrin, “XGBoost: A scalable tree boosting system,” in Proc. 22nd ACM SIGKDD Int. Conf. Knowledge Discovery and Data Mining, 2016, pp. 785–794. [Google Scholar] [Crossref]
8. D. Chicco and G. Jurman, “The advantages of the Matthews correlation coefficient (MCC) over F1 score and accuracy in binary classification evaluation,” Bioinformatics, vol. 36, no. 10, pp. 3341–3342, 2020. [Google Scholar] [Crossref]
9. A. Fernández, S. García, F. Herrera, and N. V. Chawla, “SMOTE for learning from imbalanced data: Progress and challenges,” IEEE Transactions on Knowledge and Data Engineering, 2020. [Google Scholar] [Crossref]
10. U. Fiore, A. De Santis, F. Perla, P. Zanetti, and F. Palmieri, “Using generative adversarial networks for improving classification effectiveness in credit card fraud detection,” Information Sciences, vol. 479, pp. 448–455, 2020. [Google Scholar] [Crossref]
11. M. Islam, M. Rahman, and S. Hossain, “Interpretable fraud detection using SHAP and ensemble learning,” Expert Systems with Applications, vol. 213, pp. 118–130, 2023. [Google Scholar] [Crossref]
12. W. Jiang, X. Li, and Y. Chen, “Comparative analysis of boosting algorithms for fraud detection,” Applied Soft Computing, vol. 132, pp. 109–120, 2023. [Google Scholar] [Crossref]
13. J. Jurgovsky, M. Granitzer, K. Ziegler et al., “Sequence classification for credit-card fraud detection,” Expert Systems with Applications, vol. 100, pp. 234–245, 2018. [Google Scholar] [Crossref]
14. P. Kumar, R. Singh, and A. Sharma, “Hybrid ensemble model for fraud detection using SMOTE,” IEEE Transactions on Artificial Intelligence, 2024. [Google Scholar] [Crossref]
15. S. Kumar and P. Sharma, “Machine learning approaches for fraud detection: A comparative analysis,” International Journal of Computer Applications, vol. 176, no. 32, pp. 1–7, 2020. [Google Scholar] [Crossref]
16. V. Kumar et al., “Fraud detection in financial transactions: A review,” IEEE Systems Journal, 2019. [Google Scholar] [Crossref]
17. S. M. Lundberg and S. I. Lee, “A unified approach to interpreting model predictions,” in Proc. 31st Conf. Neural Information Processing Systems (NeurIPS), 2017, pp. 4765–4774. [Google Scholar] [Crossref]
18. S. M. Lundberg, G. Erion, H. Chen et al., “From local explanations to global understanding with explainable AI for trees,” Nature Machine Intelligence, vol. 2, pp. 252–263, 2020. [Google Scholar] [Crossref]
19. T. Nguyen, T. Pham, and H. Tran, “Machine learning techniques for financial fraud detection: A survey,” IEEE Access, vol. 9, pp. 123456–123470, 2021. [Google Scholar] [Crossref]
20. M. T. Ribeiro, S. Singh, and C. Guestrin, “Why should I trust you? Explaining the predictions of any classifier,” in Proc. 22nd ACM SIGKDD Int. Conf. Knowledge Discovery and Data Mining, 2016, pp. 1135–1144. [Google Scholar] [Crossref]
21. W. Samek, G. Montavon, S. Lapuschkin, C. Anders, and K. R. Müller, “Explaining deep neural networks and beyond: A review of methods and applications,” Proceedings of the IEEE, vol. 109, no. 3, pp. 247–278, 2021. [Google Scholar] [Crossref]
22. J. Smith, “Global trends in financial fraud and digital risk management,” Journal of Financial Crime, vol. 27, no. 3, pp. 789–804, 2020. [Google Scholar] [Crossref]
23. Y. Zhang, S. Wang, and P. Phillips, “Financial fraud detection using Random Forest and deep learning,” Future Generation Computer Systems, vol. 117, pp. 378–391, 2021. [Google Scholar] [Crossref]
24. X. Zhang, H. Li, and Y. Wang, “Security vulnerabilities in mobile payment systems: Fraud risks and prevention,” Computers & Security, vol. 112, pp. 102–118, 2022. [Google Scholar] [Crossref]