Separating facts from fiction: Linguistic models to classify suspicious and trusted news posts on twitter

Svitlana Volkova; Kyle Shaffer; Jin Yea Jang; Nathan Hodas

Conference ProceedingsOPEN ACCESS

Separating facts from fiction: Linguistic models to classify suspicious and trusted news posts on twitter

ACL 2017 - 55th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (Long Papers) (2017) 2 647-653

DOI: 10.18653/v1/P17-2102

246Citations

371Readers

Abstract

Pew research polls report 62 percent of U.S. adults get news on social media (Gottfried and Shearer, 2016). In a December poll, 64 percent of U.S. adults said that “made-up news” has caused a “great deal of confusion” about the facts of current events (Barthel et al., 2016). Fabricated stories in social media, ranging from deliberate propaganda to hoaxes and satire, contributes to this confusion in addition to having serious effects on global stability. In this work we build predictive models to classify 130 thousand news posts as suspicious or verified, and predict four subtypes of suspicious news – satire, hoaxes, clickbait and propaganda. We show that neural network models trained on tweet content and social network interactions outperform lexical models. Unlike previous work on deception detection, we find that adding syntax and grammar features to our models does not improve performance. Incorporating linguistic features improves classification results, however, social interaction features are most informative for finer-grained separation between four types of suspicious news posts.

Cite

CITATION STYLE

APA

Volkova, S., Shaffer, K., Jang, J. Y., & Hodas, N. (2017). Separating facts from fiction: Linguistic models to classify suspicious and trusted news posts on twitter. In ACL 2017 - 55th Annual Meeting of the Association for Computational Linguistics, Proceedings of the Conference (Long Papers) (Vol. 2, pp. 647–653). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/P17-2102

Separating facts from fiction: Linguistic models to classify suspicious and trusted news posts on twitter

Abstract

Cite

Register to see more suggestions