Comparison of Feature Selection and Feature Extraction Role in Dimensionality Reduction of Big Data

  • Haidar Khalid Malik
  • Nashaat Jasim Al-Anber
N/ACitations
Citations of this article
13Readers
Mendeley users who have this article in their library.

Abstract

Recently, researchers intensified their efforts on a dataset with a large number of features named Big Data because of the technological revolution and the development in the data science sector. Dimensionality reduction technology has efficient, effective, and influential methods for analyzing this data, which contains many variables. The importance of Dimensionality Reduction technology lies in several fields, including “data processing, patterns recognition, machine learning, and data mining”. This paper compares two essential methods of dimensionality reduction, Feature Extraction and Feature Selection Which Machine Learning models frequently employ. We applied many classifiers like (Support vector machines, k-nearest neighbors, Decision tree, and Naive Bayes ) to the data of the anthropometric survey of US Army personnel (ANSUR 2) to classify the data and test the relevance of features by predicting a specific feature in USA Army personnel results showing that (k-nearest neighbors) achieved high accuracy (83%) in prediction, then reducing the dimensions by several techniques like (Highly Correlated Filter, Recursive  Feature Elimination, and principal components Analysis) results showing that (Recursive  Feature Elimination) have the best accuracy by (66%), From these results, it is clear that the efficiency of dimension reduction techniques varies according to the nature of the data. Some techniques are more efficient than others in text data and others are more efficient in dealing with images.

Cite

CITATION STYLE

APA

Haidar Khalid Malik, & Nashaat Jasim Al-Anber. (2023). Comparison of Feature Selection and Feature Extraction Role in Dimensionality Reduction of Big Data. Journal of Techniques, 5(1), 16–24. https://doi.org/10.51173/jt.v5i1.1027

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free