HATEPRISM: Policies, Platforms, and Research Integration. Advancing NLP for Hate Speech Proactive Mitigation

3Citations
Citations of this article
10Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Despite regulations imposed by nations and social media platforms, e.g. (Government of India, 2021; European Parliament and Council of the European Union, 2022), inter alia, hateful content persists as a significant challenge. Existing approaches primarily rely on reactive measures such as blocking or suspending offensive messages, with emerging strategies focusing on proactive measurements like detoxification and counterspeech. In our work, which we call HATEPRISM, we conduct a comprehensive examination of hate speech regulations and strategies from three perspectives: country regulations, social platform policies, and NLP research datasets. Our findings reveal significant inconsistencies in hate speech definitions and moderation practices across jurisdictions and platforms, alongside a lack of alignment with research efforts. Based on these insights, we suggest ideas and research direction for further exploration of a unified framework for automated hate speech moderation incorporating diverse strategies.

Cite

CITATION STYLE

APA

Rizwan, N., Yimam, S. M., Dementieva, D., Skupin, F., Fischer, T., Moskovskiy, D., … Mukherjee, A. (2025). HATEPRISM: Policies, Platforms, and Research Integration. Advancing NLP for Hate Speech Proactive Mitigation. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 16008–16022). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2025.findings-acl.824

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free