SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification

Ruth Talbot; Chloe Acheampong; Richard Wicentowski

Conference ProceedingsOPEN ACCESS

SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification

SemEval 2015 - 9th International Workshop on Semantic Evaluation, co-located with the 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2015 - Proceedings (2015) 626-630

DOI: 10.18653/v1/s15-2104

10Citations

97Readers

Abstract

This paper describes a sentiment classification system designed for SemEval-2015, Task 10, Subtask B. The system employs a constrained, supervised text categorization approach. Firstly, since thorough preprocessing of tweet data was shown to be effective in previous SemEval sentiment classification tasks, various preprocessessing steps were introduced to enhance the quality of lexical information. Secondly, a Naive Bayes classifier is used to detect tweet sentiment. The classifier is trained only on the training data provided by the task organizers. The system makes use of external human-generated lists of positive and negative words at several steps throughout classification. The system produced an overall F-score of 59.26 on the official test set.

Cite

CITATION STYLE

APA

Talbot, R., Acheampong, C., & Wicentowski, R. (2015). SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification. In SemEval 2015 - 9th International Workshop on Semantic Evaluation, co-located with the 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2015 - Proceedings (pp. 626–630). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/s15-2104

SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification

Abstract

Cite

Register to see more suggestions