Hindi-english hate speech detection: Author profiling, debiasing, and practical perspectives

44Citations
Citations of this article
50Readers
Mendeley users who have this article in their library.

Abstract

Code-switching in linguistically diverse, low resource languages is often semantically complex and lacks sophisticated methodologies that can be applied to real-world data for precisely detecting hate speech. In an attempt to bridge this gap, we introduce a three-tier pipeline that employs profanity modeling, deep graph embeddings, and author profiling to retrieve instances of hate speech in Hindi-English code-switched language (Hinglish) on social media platforms like Twitter. Through extensive comparison against several baselines on two real-world datasets, we demonstrate how targeted hate embeddings combined with social network-based features outperform state of the art, both quantitatively and qualitatively. Additionally, we present an expert-in-the-loop algorithm for bias elimination in the proposed model pipeline and study the prevalence and performance impact of the debiasing. Finally, we discuss the computational, practical, ethical, and reproducibility aspects of the deployment of our pipeline across the Web.

Cite

CITATION STYLE

APA

Chopra, S., Sawhney, R., Mathur, P., & Shah, R. R. (2020). Hindi-english hate speech detection: Author profiling, debiasing, and practical perspectives. In AAAI 2020 - 34th AAAI Conference on Artificial Intelligence (pp. 386–393). AAAI press. https://doi.org/10.1609/aaai.v34i01.5374

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free