Design of big data anomaly detection model based on random forest algorithm
DOI:
https://doi.org/10.59782/sidr.v1i1.40Keywords:
big data clustering, feature extraction, principal component analysis, random forest classifier, decision tree, update weightAbstract
Aiming at the problem that the anomaly detection process of big data is easily disturbed by edge data, resulting in poor accuracy of big data anomaly detection, a big data anomaly detection model based on random forest algorithm is proposed. Firstly, the improved -means algorithm is used to cluster the big data, and the principal component analysis method is used to extract the big data features; then, a big data anomaly detection model based on random forest classifier is constructed, the extracted features are input into the model, a decision tree is constructed, and the classification accuracy of the classifier is improved by dynamically updating the weight value of the decision tree; finally, the classification result is output to complete the anomaly detection of big data. The experimental results show that the detection time of the proposed model is about 25 s, the average accuracy of big data anomaly detection is, and the false alarm rate is .
Downloads
How to Cite
Issue
Section
License
Copyright (c) 2024 Scientific Insights and Discoveries Review

This work is licensed under a Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License.