A Pipeline for Rapid Post-Crisis Twitter Data Acquisition, Filtering and Visualization

10Citations
Citations of this article
36Readers
Mendeley users who have this article in their library.

Abstract

Due to instant availability of data on social media platforms like Twitter, and advances in machine learning and data management technology, real-time crisis informatics has emerged as a prolific research area in the last decade. Although several benchmarks are now available, especially on portals like CrisisLex, an important, practical problem that has not been addressed thus far is the rapid acquisition, benchmarking and visual exploration of data from free, publicly available streams like the Twitter API in the immediate aftermath of a crisis. In this paper, we present such a pipeline for facilitating immediate post-crisis data collection, curation and relevance filtering from the Twitter API. The pipeline is minimally supervised, alleviating the need for feature engineering by including a judicious mix of data preprocessing and fast text embeddings, along with an active learning framework. We illustrate the utility of the pipeline by describing a recent case study wherein it was used to collect and analyze millions of tweets in the immediate aftermath of the Las Vegas shootings in 2017.

Cite

CITATION STYLE

APA

Kejriwal, M., & Gu, Y. (2019). A Pipeline for Rapid Post-Crisis Twitter Data Acquisition, Filtering and Visualization. Technologies, 7(2). https://doi.org/10.3390/technologies7020033

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free