SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification

10Citations
Citations of this article
97Readers
Mendeley users who have this article in their library.

Abstract

This paper describes a sentiment classification system designed for SemEval-2015, Task 10, Subtask B. The system employs a constrained, supervised text categorization approach. Firstly, since thorough preprocessing of tweet data was shown to be effective in previous SemEval sentiment classification tasks, various preprocessessing steps were introduced to enhance the quality of lexical information. Secondly, a Naive Bayes classifier is used to detect tweet sentiment. The classifier is trained only on the training data provided by the task organizers. The system makes use of external human-generated lists of positive and negative words at several steps throughout classification. The system produced an overall F-score of 59.26 on the official test set.

Cite

CITATION STYLE

APA

Talbot, R., Acheampong, C., & Wicentowski, R. (2015). SWASH: A Naive Bayes Classifier for Tweet Sentiment Identification. In SemEval 2015 - 9th International Workshop on Semantic Evaluation, co-located with the 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2015 - Proceedings (pp. 626–630). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/s15-2104

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free