NEARMISS-SF ALGORITHM FOR DEALING WITH HIGH – DIMENSIONAL IMBALANCED DATA SETS
Abstract
Tóm tắt
Article Details

This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.
References
He H. and Garcia E. A. (2009). Learning from Imbalanced Data. IEEE Trans. Knowl. Data Eng. 21 (9): 1263–1284.
Chawla N. V, Bowyer K. W, Hall L. O., and Kegelmeyer W. P. (2002). SMOTE: Synthetic Minority Over-sampling Technique. Artif. Intell. Res. 16 (1): 321–357.
Han H., Wang W.Y., and Mao B.H. (2005). Borderline-SMOTE: A New Over-Sampling Method in Imbalanced Data Sets Learning. Advances in Intelligent Computing, ICIC 2005, Lecture Notes in Computer Science. 3644
Bunkhumpornpat C., Sinapiromsaran K., and Lursinsap C. (2009). Safe-Level-SMOTE: Safe-Level-Synthetic Minority Over-sampling Technique for handling the class imbalanced problem. 5476: 475–482
Maldonado S., López J., Vairetti C. (2018). An alternative SMOTE oversampling strategy for high-dimensional datasets. Applied Soft Computing Journal. 76: 380–389
Zhang Z. and Mani I. (2003). KNN Approach to Unbalanced Data Distribution: A Case Study involving Information Extraction. Workshop on Learning from Inmbalanced Datasets II, ICML, Washington DC
Sun Y., Wong A. K. C., and Kamel M. S. (2009). Classification of Imbalanced Data: A Review. J. Pattern Recognit. 23 (4): 687–719
Lichman M. (2013). UCI Machine Learning Repository, http://archive.ics.uci.edu/ml, Irvine, CA: University of California, School of Information and Computer Science.
Hall M., Frank E., Holmes G., Pfahringer B., Reutemann P., Written L. (2009). The WEKA Data Mining Software: An Update. ACM SIGKDD Explorations Newsletter. 11 (1): 10-18.
[10] Gower J.C. (1971). “A general coefficient of similarity and some of its properties”. Biometrics. 27 (1): 857–874