Repurposing benchmark corpora for reconstructing provenance

Sara Magliacane; Paul Groth

Conference Proceedings

Repurposing benchmark corpora for reconstructing provenance

CEUR Workshop Proceedings (2013) 994 39-50

ISSN: 16130073

1Citations

11Readers

Abstract

Provenance is a critical aspect in evaluating scientific output, yet, it is still often overlooked or not comprehensively produced by practitioners. This incomplete and partial nature of provenance has been recognized in the literature, which has led to the development of new methods for reconstructing missing provenance. Unfortunately, there is currently no agreed upon evaluation framework for testing these methods. Moreover, there is a paucity of datasets that these methods can be applied to. To begin to address this gap, we present a survey of existing benchmark corpora from other computer science communities that could be applied to evaluate provenance reconstruction techniques. The survey identifies, for each corpus, a mapping between the data available and common provenance concepts. In addition to their applicability to provenance reconstruction, we also argue that these corpora could be reused for other tasks pertaining to provenance.

Cite

CITATION STYLE

APA

Magliacane, S., & Groth, P. (2013). Repurposing benchmark corpora for reconstructing provenance. In CEUR Workshop Proceedings (Vol. 994, pp. 39–50). CEUR-WS.

Repurposing benchmark corpora for reconstructing provenance

Abstract

Cite

Register to see more suggestions