Evaluating hypotheses in geolocation on a very large sample of Twitter

1Citations
Citations of this article
72Readers
Mendeley users who have this article in their library.

Abstract

Recent work in geolocation has made several hypotheses about what linguistic markers are relevant to detect where people write from. In this paper, we examine six hypotheses against a corpus consisting of all geo-tagged tweets from the US, or whose geo-tags could be inferred, in a 19% sample of Twitter history. Our experiments lend support to all six hypotheses, including that spelling variants and hashtags are strong predictors of location. We also study what kinds of common nouns are predictive of location after controlling for named entities such as dolphins or sharks.

Cite

CITATION STYLE

APA

Salehi, B., & Søgaard, A. (2017). Evaluating hypotheses in geolocation on a very large sample of Twitter. In 3rd Workshop on Noisy User-Generated Text, W-NUT 2017 - Proceedings of the Workshop (pp. 62–67). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/w17-4409

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free