A semi-supervised active-learning truth estimator for social networks

10Citations
Citations of this article
26Readers
Mendeley users who have this article in their library.
Get full text

Abstract

This paper introduces an active-learning-based truth estimator for social networks, such as Twitter, that enhances estimation accuracy significantly by requesting a well-selected (small) fraction of data to be labeled. Data assessment and truth discovery from arbitrary open online sources are a hard problem due to uncertainty regarding source reliability. Multiple truth finding systems were developed to solve this problem. Their accuracy is limited by the noisy nature of the data, where distortions, fabrications, omissions, and duplication are introduced. This paper presents a semi-supervised truth estimator for social networks, in which a portion of inputs are carefully selected to be reliably verified. The challenge is to find the subset of observations to verify that would maximally enhance the overall fact-finding accuracy. This work extends previous passive approaches to recursive truth estimation, as well as semi-supervised approaches where the estimator has no control over the choice of data to be labeled. Results show that by optimally selecting claims to be verified, we improve estimated accuracy by 12% over unsupervised baseline, and by 5% over previous semi-supervised approaches.

Cite

CITATION STYLE

APA

Cui, H., Abdelzaher, T., & Kaplan, L. (2019). A semi-supervised active-learning truth estimator for social networks. In The Web Conference 2019 - Proceedings of the World Wide Web Conference, WWW 2019 (pp. 296–306). Association for Computing Machinery, Inc. https://doi.org/10.1145/3308558.3313712

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free