Abstract
This paper discusses the question whether it is possible to learn a generic representation that is useful for detecting various types of abusive language. The approach is inspired by recent advances in transfer learning and word embeddings, and we learn representations from two different datasets containing various degrees of abusive language. We compare the learned representation with two standard approaches; one based on lexica, and one based on data-specific n-grams. Our experiments show that learned representations do contain useful information that can be used to improve detection performance when training data is limited.
Cite
CITATION STYLE
Sahlgren, M., Isbister, T., & Olsson, F. (2018). Learning Representations for Detecting Abusive Language. In 2nd Workshop on Abusive Language Online - Proceedings of the Workshop, co-located with EMNLP 2018 (pp. 115–123). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/w18-5115
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.