Improving fraud prediction with incremental data balancing technique for massive data streams
The performance of classification algorithms with a massive and highly imbalanced data stream depends upon efficient balancing strategy. Some techniques of balancing strategy have been applied in the past with Batch data to resolve the class imbalance problem. This paper proposes a new incremental data balancing framework which can work with massive imbalanced data streams. In this paper, we choose Racing Algorithm as an automated data balancing technique which optimizes the balancing techniques. We applied Random Forest classification algorithm which can deal with the massive data stream. We investigated the suitability of Racing Algorithm and Random Forest in the proposed framework. Applying new technique in the proposed framework on the European Credit Card dataset, provided better results than the Batch mode. The proposed framework is more scalable to handle online massive data streams.
Code (0)
등록된 구현이 없습니다.
Tasks
General ClassificationSimilar Papers 제목 키워드 기반
A Data Balancing and Ensemble Learning Approach for Credit Card Fraud Detection
This research introduces an innovative method for identifying credit card fraud by combining the SMOTE-KMEANS technique with an ensemble machine learning model. The proposed model was benchmarked against traditional mode…
Ensemble LearningFraud DetectionDynamic Fraud Detection: Integrating Reinforcement Learning into Graph Neural Networks
Financial fraud refers to the act of obtaining financial benefits through dishonest means. Such behavior not only disrupts the order of the financial market but also harms economic and social development and breeds other…
Fraud DetectionIncremental Outlier Detection Modelling Using Streaming Analytics in Finance & Health Care
In this paper, we had built the online model which are built incrementally by using online outlier detection algorithms under the streaming environment. We identified that there is highly necessity to have the streaming …
Diabetes PredictionFraud DetectionOutlier DetectionPredictionInstance-Level Explanations for Fraud Detection: A Case Study
Fraud detection is a difficult problem that can benefit from predictive modeling. However, the verification of a prediction is challenging; for a single insurance policy, the model only provides a prediction score. We pr…
Fraud DetectionPredictionAn Efficient Machine Learning-based Framework for Detection and Prevention of Frauds in Telecom Networks
Telecommunication fraud is an acute problem that leads to substantial material losses and compromises the reliability of telecom systems worldwide. Only effective and efficient detection mechanisms can help to deal with …
Fraud Detection