A New Hate Speech Detection System based on Textual and Psychological Features

9Citations
Citations of this article
42Readers
Mendeley users who have this article in their library.

Abstract

Hate speech often spreads on social media and harms individuals and the community. Machine learning models have been proposed to detect hate speech in social media; however, several issues presently limit the performance of current approaches. One challenge is the issue of having diverse comprehensions of hate speech constructs which will lead to many speech categories and different interpretations. In addition, certain language-specific features, and short text issues, such as Twitter, exacerbate the problem. Moreover, current machine learning approaches lack universality due to small datasets and the adoption of a few features of hateful speech. This paper develops and builds new feature sets based on frequencies of textual tokens and psychological characteristics. Then, the study evaluates several machine learning methods over a large dataset. Results showed that the Random Forest and BERT methods are the most valuable for detecting hate speech content. Furthermore, the most dominant features that are helpful for hate speech detection methods combine psychological features and Term-Frequency Inverse Document-Frequency (TFIDF) features. Therefore, the proposed approach could identify hate speech on social media platforms like Twitter.

Cite

CITATION STYLE

APA

Alkomah, F., Salati, S., & Ma, X. (2022). A New Hate Speech Detection System based on Textual and Psychological Features. International Journal of Advanced Computer Science and Applications, 13(8), 860–869. https://doi.org/10.14569/IJACSA.2022.01308100

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free