Abstract
The structure and behavior of human networks have been investigated and quantitatively modeled by modern social scientists for decades, however the scope of these efforts is often constrained by the labor-intensive curation processes that are required to collect, organize, and analyze network data. The surge in online social media in recent years provides a new source of dynamic, semi-structured data of digital human networks, many of which embody attributes of real-world networks. In this paper we leverage the Reddit social media platform to study social communities whose dynamics indicate they may have experienced a disturbance event. We describe an unsupervised approach to analyzing natural language content for quantifying community similarity, monitoring temporal changes, and detecting anomalies indicative of disturbance events. We demonstrate how this method is able to detect anomalies in a spectrum of Reddit communities and discuss its applicability to unsupervised event detection for a broader class of social media use cases.
Cite
CITATION STYLE
Shah, D., Hurley, M., Liu, J., & Daggett, M. (2019). Unsupervised content-based characterization and anomaly detection of online community dynamics. In Proceedings of the Annual Hawaii International Conference on System Sciences (Vol. 2019-January, pp. 2264–2273). IEEE Computer Society. https://doi.org/10.24251/hicss.2019.274
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.