Sentiment analysis on monolingual, multilingual and code-switching twitter corpora

63Citations
Citations of this article
124Readers
Mendeley users who have this article in their library.

Abstract

We address the problem of performing polarity classification on Twitter over different languages, focusing on English and Spanish, comparing three techniques: (1) a monolingual model which knows the language in which the opinion is written, (2) a monolingual model that acts based on the decision provided by a language identification tool and (3) a multilingual model trained on a multilingual dataset that does not need any language recognition step. Results show that multilingual models are even able to outperform the monolingual models on some monolingual sets. We introduce the first code-switching corpus with sentiment labels, showing the robustness of a multilingual approach.

Cite

CITATION STYLE

APA

Vilares, D., Alonso, M. A., & Gómez-Rodríguez, C. (2015). Sentiment analysis on monolingual, multilingual and code-switching twitter corpora. In 6th Workshop on Computational Approaches to Subjectivity, Sentiment and Social Media Analysis, WASSA 2015 at the 2015 Conference on Empirical Methods in Natural Language Processing, EMNLP 2015 - Proceedings (pp. 2–8). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/w15-2902

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free