Classification of imbalanced data using support vector machine and rough set theory: A review

17Citations
Citations of this article
27Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

The performance of machine learning classifier such as support vector machine (SVM) degraded by the nature and structural construct of real-world data which is in most cases are imbalanced. The accuracy and decision making typically biased towards majority class and this significantly affect the result of the classification of minority class. Nevertheless, dataset does not always comprise of significant attributes even with large number of points in certain class, but rather it could potentially lead to redundancy and irrelevant features. Rough set (RS) theory is a mathematical tool for tackling ambiguity and removing redundancy in the dataset. This can further help the classification system in improving its accuracy of the prediction for both majority and minority class. Commonly, RS theory was utilised as a preprocessing method to bring about the knowledge, association rules, or potential patterns in the data. The output of RS theory is a reduced set of attributes which contains same indiscernibility as the original dataset. Hence, the focus of this paper is a review of literature and findings on the classification strategy which employs SVM and RS as a combined system to solve the problem of imbalanced data.

Cite

CITATION STYLE

APA

Ibrahim, H., Anwar, S. A., & Ahmad, M. I. (2021). Classification of imbalanced data using support vector machine and rough set theory: A review. In Journal of Physics: Conference Series (Vol. 1878). IOP Publishing Ltd. https://doi.org/10.1088/1742-6596/1878/1/012054

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free