A Theoretical Framework and 38-Feature Taxonomy for Behavioral Fraud Detection in Digital Banking
DOI:
https://doi.org/10.59261/jequi.v8i3.360Keywords:
Fraud Detection, Machine Learning, CatBoost, LightGBM, Risk Assessment, Fintech, Behavioural ProfilingAbstract
Background: Traditional rule-based systems are increasingly inadequate for addressing dynamic financial threats because of their reactive nature and susceptibility to concept drift. At PT XYZ, this vulnerability resulted in a sharp decline in fraud prevention effectiveness, from 90.4% in 2022 to 75.3% in 2025, leaving a critical 24.7% security gap.
Objective: This study aimed to develop a behavioral fraud detection framework for digital banking by integrating a 38-feature account-level taxonomy with Isolation Forest, LightGBM, CatBoost, and the Synthetic Minority Over-sampling Technique (SMOTE) to overcome the limitations of legacy rule-based systems.
Methods: The study established a proactive theoretical framework using the Cross-Industry Standard Process for Data Mining (CRISP-DM) methodology to shift from transaction-level monitoring to holistic account-level behavioral profiling. Central to this approach was a 38-feature taxonomy that integrated statistical aggregates, temporal signatures, velocity flags, and volume metrics to map the behavioral “fingerprint” of at-risk accounts.
Results: Empirical testing on 402 unique accounts confirmed that the proposed architecture successfully closed the identified 24.7% security gap. Both the LightGBM and CatBoost classifiers achieved a perfect recall score of 1.0000 and an accuracy of 0.9972. LightGBM emerged as the best-performing model, with a receiver operating characteristic area under the curve (ROC AUC) of 0.9996, significantly outperforming traditional logistic regression.
Conclusion: The proposed framework effectively improved fraud detection by combining behavioral profiling, hybrid anomaly detection, and balanced ensemble learning. LightGBM and CatBoost achieved 99.72% accuracy, 100% recall, and a 99.96% ROC AUC, demonstrating a practical and explainable approach to proactive digital banking fraud detection.
Downloads
References
Ali, A., Razak, S., Othman, S., Eisa, T., Al-Dhaqm, A., Nasser, M., Elhassan, T., Elshafie, H., & Saif, A. (2022). Financial Fraud Detection Based on Machine Learning: A Systematic Literature Review. Applied Sciences, 12(19), 9637. https://doi.org/10.3390/app12199637
Balcioglu, Y., Merter, A., & Çelik, B. (2025). Behavioral Analysis of Customer Transaction Patterns In Financial Fraud Detection. https://www.researchgate.net/publication/384567123
Brause, R., Langsdorf, T., & Hepp, M. (1999). Neural data mining for credit card fraud detection. In Proceedings of the 11th IEEE International Conference on Tools with Artificial Intelligence (pp. 103–106). https://doi.org/10.1109/TAI.1999.809773
Carcillo, F., Le Borgne, Y. A., Caelen, O., Oblé, F., & Bontempi, G. (2021). Combining unsupervised And Supervised Learning in Credit Card Fraud Detection. Information Sciences, 557, 317–331. https://doi.org/10.1016/j.ins.2019.05.042
Farahani, R. G., Zarrabi, A., & Ghazanfari, P. (2025). A Report on CatBoost: Unbiased Boosting with Categorical Features. Accessed: Aug, 11.
Fatlawi, H. K. (2025). Enhanced Fraudulent Detection Using Isolation Forest and Multi-Cluster Deep Learning. Journal of Al-Qadisiyah for Computer Science and Mathematics, 17(1), 45–58.
Fiore, U., De Santis, A., Perla, F., Zanetti, P., & Palmieri, F. (2021). Using Generative Adversarial Networks for Improving Classification Effectiveness in Credit Card Fraud Detection. Applied Soft Computing, 112, 107793. https://doi.org/10.1016/j.asoc.2021.107793
Hilal, W., Gadsden, S. A., & Yawney, J. (2022). Financial Fraud: A review of anomaly detection techniques and recent advances. Expert Systems with Applications, 193, 116429. https://doi.org/10.1016/j.eswa.2021.116429
Hindarto, D. (2024). Case Study: Gradient Boosting Machine VS Light GBM in Potential Landslide Detection. Journal of Computer Networks, Architecture and High Performance Computing, 6(1), 169–178.
Kalid, S. N., Khor, K. C., Ng, K. H., & Tong, G. K. (2024). Detecting frauds and Payment Defaults on Credit Card Data Inherited with Imbalanced Class Distribution and Overlapping Class Problems: A Systematic Review. IEEE Access, 12, 23636–23652.
Ke, G., Meng, Q., Finley, T., Wang, T., Chen, W., Ma, W., Ye, Q., & Liu, T.-Y. (2017). LightGBM: A highly Efficient Gradient Boosting Decision Tree. In Advances in Neural Information Processing Systems (Vol. 30, pp. 3146–3154). https://proceedings.neurips.cc/paper_files/paper/2017/hash/6449f44a102fde848669bdd9eb6b76fa-Abstract.html
Keating, L. (2023). Concept Drift Handling in Continuous Fraud Detection Systems. Journal of Financial Cybersecurity, 5(2), 78–94.
Mienye, I. D., & Swart, T. G. (2024). A Hybrid Deep Learning Approach with Generative Adversarial Network For Credit Card Fraud Detection. Technologies, 12(10), 186.
Pourhabibi, T. (2024). Fraud Detection: A Graph Based Anomaly Detection Approach. RMIT University.
Pourhabibi, T., Ong, K. L., Kam, B. H., & Boo, Y. L. (2020). Fraud Detection: A Systematic Literature Review of Graph-Based Anomaly Detection Approaches. Decision Support Systems, 133, 113303. https://doi.org/10.1016/j.dss.2020.113303
Prokhorenkova, L., Gusev, G., Vorobev, A., Dorogush, A. V, & Gulin, A. (2018). CatBoost: Unbiased Boosting with Categorical Features. In Advances in Neural Information Processing Systems (Vol. 31, pp. 6638–6648). https://proceedings.neurips.cc/paper_files/paper/2018/hash/14491b756b3a51daac41c24863285549-Abstract.html
Taha, A. A., & Malebary, S. J. (2020). An Intelligent Approach to Credit Card Fraud Detection Using An Optimized Light Gradient Boosting Machine. IEEE Access, 8, 25579–25587. https://doi.org/10.1109/ACCESS.2020.2971354
Wang, H., Zhu, Y., & Adam, H. (2023). AutoEncoder and LightGBM for Credit Card Fraud Detection Problems. Symmetry, 15(4), 870. https://doi.org/10.3390/sym15040870
Wang, Y., Xu, W., & Zhang, H. (2023). Adaptive Fraud Detection Systems: A Review of Machine Learning Approaches Under Concept Drift. IEEE Transactions on Knowledge and Data Engineering, 35(4), 3892–3910. https://doi.org/10.1109/TKDE.2022.3184842
Xiao, Y., Tan, L., & Liu, J. (2025). Application of Machine Learning Model in Fraud Identification: A Comparative Study of CatBoost, XGBoost and LightGBM. Preprints. https://doi.org/10.20944/preprints202503.1199.v1
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Ardi Darmawan, Thoyyibah T

This work is licensed under a Creative Commons Attribution-ShareAlike 4.0 International License.
Authors who publish with this journal agree to the following terms:
- Authors retain copyright and grant the journal right of first publication with the work simultaneously licensed under a Creative Commons Attribution-ShareAlike 4.0 International (CC-BY-SA). that allows others to share the work with an acknowledgement of the work's authorship and initial publication in this journal.
- Authors are able to enter into separate, additional contractual arrangements for the non-exclusive distribution of the journal's published version of the work (e.g., post it to an institutional repository or publish it in a book), with an acknowledgement of its initial publication in this journal.
- Authors are permitted and encouraged to post their work online (e.g., in institutional repositories or on their website) prior to and during the submission process, as it can lead to productive exchanges, as well as earlier and greater citation of published work.



