Effectiveness of SMOTE-ENN to Reduce Complexity in Classification Model

Main Article Content

Ines Riantika
Bagus Sartono
Khairil Anwar Notodiputro

Abstract

A failure to produce classification models with high performance might be caused by the dataset's characteristics, such as the between-class overlapping and the class imbalance. The higher the data complexity, the more complicated it is for the algorithm to find good models.  Combining the issues of class imbalance and overlapping would make the problem more challenging. To deal with this problem, this research implemented a hybrid class-balancing technique named SMOTE-ENN. This technique adds observations to the minority class to balance the class frequencies.  After that, it removes some observations to reduce the degree of overlapping.  The research revealed that SMOTE-ENN succeeds in doing that.  We employed a random forest method to evaluate it. In 28 out of 46 cases we investigated, the new datasets generated by SMOTE-ENN could produce models with higher accuracy.

Downloads

Download data is not yet available.

Article Details

How to Cite
1.
Riantika I, Sartono B, Anwar Notodiputro K. Effectiveness of SMOTE-ENN to Reduce Complexity in Classification Model. IJSA [Internet]. 2024 Jun. 11 [cited 2025 Nov. 29];8(1):70-82. Available from: https://journal-stats.ipb.ac.id/index.php/ijsa/article/view/1203
Section
Articles

Most read articles by the same author(s)