Contextual Word Embedding: A Case Study in Clustering Tweets about Emergency Situations

6Citations
Citations of this article
12Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Effective clustering of short documents, such as tweets, is difficult because of the lack of sufficient semantic context. Word embedding is a technique that is effective in addressing this lack of semantic context. However, the process of word vector embedding, in turn, relies on the availability of sufficient contexts to learn the word associations. To get around this problem, we propose a novel word vector training approach that leverages topically similar tweets to better learn the word associations. We test our proposed word embedding approach by clustering a collection of tweets on disasters. We observe that the proposed method improves clustering effectiveness by up to 14%.

Cite

CITATION STYLE

APA

Ganguly, D., & Ghosh, K. (2018). Contextual Word Embedding: A Case Study in Clustering Tweets about Emergency Situations. In The Web Conference 2018 - Companion of the World Wide Web Conference, WWW 2018 (pp. 73–74). Association for Computing Machinery, Inc. https://doi.org/10.1145/3184558.3186935

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free