CASTLE: Crowd-Assisted System for Text Labeling and Extraction

8Citations
Citations of this article
7Readers
Mendeley users who have this article in their library.

Abstract

The amount of text data has been growing exponentially and with it the demand for improved information extraction (IE) efforts to analyze and query such data. While automatic IE systems have proven useful in controlled experiments, in practice the gap between machine learning extraction and human extraction is still quite large. In this paper, we propose a system that uses crowdsourcing techniques to help close this gap. One of the fundamental issues inherent in using a large-scale human workforce is deciding the optimal questions to pose to the crowd. We demonstrate novel solutions using mutual information and token clustering techniques in the domain of bibliographic citation extraction. Our experiments show promising results in using crowd assistance as a cost-effective way to close up the”last mile” between extraction systems and a human annotator.

Cite

CITATION STYLE

APA

Goldberg, S., Wang, D. Z., & Kraska, T. (2013). CASTLE: Crowd-Assisted System for Text Labeling and Extraction. In Proceedings of the 1st AAAI Conference on Human Computation and Crowdsourcing, HCOMP 2013 (pp. 51–59). AAAI Press. https://doi.org/10.1609/hcomp.v1i1.13087

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free