GARS: Genetic Algorithm for the identification of a Robust Subset of features in high-dimensional datasets

Mattia Chiesa; Giada Maioli; Gualtiero I. Colombo; Luca Piacentini

Journal ArticleOPEN ACCESS

GARS: Genetic Algorithm for the identification of a Robust Subset of features in high-dimensional datasets

BMC Bioinformatics (2020) 21(1)

DOI: 10.1186/s12859-020-3400-6

30Citations

69Readers

Abstract

Background: Feature selection is a crucial step in machine learning analysis. Currently, many feature selection approaches do not ensure satisfying results, in terms of accuracy and computational time, when the amount of data is huge, such as in 'Omics' datasets. Results: Here, we propose an innovative implementation of a genetic algorithm, called GARS, for fast and accurate identification of informative features in multi-class and high-dimensional datasets. In all simulations, GARS outperformed two standard filter-based and two 'wrapper' and one embedded' selection methods, showing high classification accuracies in a reasonable computational time. Conclusions: GARS proved to be a suitable tool for performing feature selection on high-dimensional data. Therefore, GARS could be adopted when standard feature selection approaches do not provide satisfactory results or when there is a huge amount of data to be analyzed.

Author supplied keywords

Cite

CITATION STYLE

APA

Chiesa, M., Maioli, G., Colombo, G. I., & Piacentini, L. (2020). GARS: Genetic Algorithm for the identification of a Robust Subset of features in high-dimensional datasets. BMC Bioinformatics, 21(1). https://doi.org/10.1186/s12859-020-3400-6

GARS: Genetic Algorithm for the identification of a Robust Subset of features in high-dimensional datasets

Abstract

Author supplied keywords

Cite

Register to see more suggestions