Text Classification of Cornell Movie Data using Data Mining with Feature Selection

undefined; undefined; undefined; A. K. Shrivas; S. M. Ghosh; Amit Kumar Dewangan

Journal Article

Text Classification of Cornell Movie Data using Data Mining with Feature Selection

et al.

International Journal of Engineering and Advanced Technology (2019) 9(2) 2950-2955

DOI: 10.35940/ijeat.b2329.129219

N/ACitations

3Readers

Get full text

Abstract

Text Classification is branch of text mining through which we can analyze the sentiment of the movie data. In this research paper we have applied different preprocessing techniques to reduce the features from cornell movie data set. We have also applied the Correlation-based feature subset selection and chi-square feature selection technique for gathering most valuable words of each category in text mining processes. The new cornell movie data set formed after applying the preprocessing steps and feature selection techniques. We have classified the cornell movie data as positive or negative using various classifiers like Support Vector Machine (SVM), Multilayer Perceptron (MLP), Naive Bayes (NB), Bays Net (BN) and Random Forest (RF) classifier. We have also compared the classification accuracy among classifiers and achieved better accuracy i. e. 87% in case of SVM classifier with reduced number of features. The suggested classifier can be useful in opinion of movie review, analysis of any blog and documents etc.

Cite

CITATION STYLE

APA

Shrivas, A. K., Ghosh, S. M., & Dewangan, A. K. (2019). Text Classification of Cornell Movie Data using Data Mining with Feature Selection. International Journal of Engineering and Advanced Technology, 9(2), 2950–2955. https://doi.org/10.35940/ijeat.b2329.129219

Text Classification of Cornell Movie Data using Data Mining with Feature Selection

Abstract

Cite

Register to see more suggestions