Provenance management in curated databases

205Citations
Citations of this article
177Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Curated databases in bioinformatics and other disciplines are the result of a great deal of manual annotation, correction and transfer of data from other sources. Provenance information concerning the creation, attribution, or version history of such data is crucial for assessing its integrity and scientific value. General purpose database systems provide little support for tracking provenance, especially when data moves among databases. This paper investigates general-purpose techniques for recording provenance for data that is copied among databases. We describe an approach in which we track the user's actions while browsing source databases and copying data into a curated database, in order to record the user's actions in a convenient, queryable form. We present an implementation of this technique and use it to evaluate the feasibility of database support for provenance management. Our experiments show that although the overhead of a naive approach is fairly high, it can be decreased to an acceptable level using simple optimizations. Copyright 2006 ACM.

Author supplied keywords

Cite

CITATION STYLE

APA

Buneman, P., Chapman, A., & Cheney, J. (2006). Provenance management in curated databases. In Proceedings of the ACM SIGMOD International Conference on Management of Data (pp. 539–550). https://doi.org/10.1145/1142473.1142534

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free