Abstract
This paper describes a simple data model for the composition and metadata management of documents in a distributed setting. We assume that each document resides at the local repository of its provider, so all providers' repositories, collectively, can be thought of as a single database of documents spread over the network. Providers willing to share their documents with other providers in the network must register them with a coordinator, or mediator, and providers that search for documents matching their needs must address their queries to the mediator. The process of registering (or un-registering) a document, formulating a query to the mediator, or answering a query by the mediator, all rely on document content annotation. Content annotation depends on the nature of the document: if the document is atomic then an annotation provided explicitely by the author is sufficient, whereas if the document is composite then the author annotation should be augmented by an implied annotation, i.e., an annotation inferred from the annotations of the document's components. The main contributions of this paper are: 1. Providing appropriate definitions of document annotations; 2. Providing an algorithm for the automatic computation of implied annotations; 3. Defining the main services that the mediator should support. © Springer-Verlag Berlin Heidelberg 2004.
Cite
CITATION STYLE
Rigaux, P., & Spyratos, N. (2004). Metadata inference for document retrieval in a distributed repository. Lecture Notes in Computer Science (Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), 3321, 418–436. https://doi.org/10.1007/978-3-540-30502-6_31
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.