SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection

25Citations
Citations of this article
30Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Hate speech detection on online social networks has become one of the emerging hot topics in recent years. With the broad spread and fast propagation speed across online social networks, hate speech makes significant impacts on society by increasing prejudice and hurting people. Therefore, there are aroused attention and concern from both industry and academia. In this paper, we address the hate speech problem and propose a novel hate speech detection framework called SWE2, which only relies on the content of messages and automatically identifies hate speech. In particular, our framework exploits both word-level semantic information and sub-word knowledge. It is intuitively persuasive and also practically performs well under a situation with/without character-level adversarial attack. Experimental results show that our proposed model achieves 0.975 accuracy and 0.953 macro F1, outperforming 7 state-of-the-art baselines under no adversarial attack. Our model robustly and significantly performed well under extreme adversarial attack (manipulation of 50% messages), achieving 0.967 accuracy and 0.934 macro F1.

Cite

CITATION STYLE

APA

Mou, G., Ye, P., & Lee, K. (2020). SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection. In International Conference on Information and Knowledge Management, Proceedings (pp. 1145–1154). Association for Computing Machinery. https://doi.org/10.1145/3340531.3411990

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free