Skip to the main content

Review article

https://doi.org/10.32985/ijeces.11.1.2

In-Depth Performance Analysis of SMOTE-Based Oversampling Algorithms in Binary Classification

Mario Dudjak orcid id orcid.org/0000-0001-9424-2251 ; J. J. Strossmayer University of Osijek, Faculty of Electrical Engineering, Computer Science and Information Technology
Goran Martinović orcid id orcid.org/0000-0002-7469-6018 ; J. J. Strossmayer University of Osijek, Faculty of Electrical Engineering, Computer Science and Information Technology


Full text: english pdf 280 Kb

page 13-23

downloads: 755

cite


Abstract

In the field of machine learning, the problem of class imbalance considerably impairs the performance of classification algorithms. Various techniques have been proposed that seek to mitigate classifier bias with respect to the majority class, with simple oversampling approaches being one of the most effective. Their main representative is the well-known SMOTE algorithm, which introduces a synthetic instances creation mechanism as an interpolation procedure between minority instances. To date, an abundance of SMOTE-based extensions that intend to improve the original algorithm has been proposed. This paper aims to compare the performance of several such extensions. In addition to comparing the overall performance, the impact of the selected oversamplers on the per-class performance is also evaluated. Finally, this paper tries to interpret the obtained performance results with respect to the internal procedures of oversampling algorithms. Some interesting findings have been made in this regard.

Keywords

classification, class imbalance, SMOTE, sensitivity, specificity

Hrčak ID:

242928

URI

https://hrcak.srce.hr/242928

Publication date:

15.4.2020.

Visits: 1.807 *