Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers

1Citations
Citations of this article
23Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The proliferation of hate speech on social media necessitates automated detection systems that balance accuracy with computational efficiency. This study evaluates 38 model configurations in detecting hate speech across datasets ranging from 6.5K to 451K samples. We analyze transformer architectures (e.g., BERT, RoBERTa, Distil-BERT), deep neural networks (e.g., CNN, LSTM, GRU, Hierarchical Attention Networks), and traditional machine learning methods (e.g., SVM, CatBoost, Random Forest).Our results show that transformers, particularly RoBERTa, consistently achieve superior performance with accuracy and F1-scores exceeding 90%. Among deep learning approaches, Hierarchical Attention Networks yield the best results, while traditional methods like CatBoost and SVM remain competitive, achieving F1-scores above 88% with significantly lower computational costs.Additionally, our analysis highlights the importance of dataset characteristics, with balanced, moderately sized unprocessed datasets outperforming larger, preprocessed datasets. These findings offer valuable insights for developing efficient and effective hate speech detection systems.

Cite

CITATION STYLE

APA

Abusaqer, M., Saquer, J., & Shatnawi, H. (2026). Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers. In Proceedings of the 2025 ACM Southeast Conference, ACMSE 2025 (pp. 203–213). Association for Computing Machinery, Inc. https://doi.org/10.1145/3696673.3723061

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free