Caching and Database Scaling in Distributed Shared-Nothing Information Retrieval Systems

16Citations
Citations of this article
11Readers
Mendeley users who have this article in their library.

Abstract

A common class of existing information retrieval system provides access to abstracts. For example Stanford University, through its FOLIO system, provides access to the INSPECT database of abstracts of the literature on physics, computer science, electrical engineering, etc. In this paper this database is studied by using a trace-driven simulation. We focus on physical index design, inverted index caching, and database scaling in a distributed shared-nothing system. All three issues are shown to have a strong effect on response time and throughput. Database scaling is explored in two ways. One way assumes an “optimal” configuration for a single host and then linearly scales the database by duplicating the host architecture as needed. The second way determines the optimal number of hosts given a fixed database size. © 1993, ACM. All rights reserved.

Cite

CITATION STYLE

APA

Tomasic, A., & Garcia-Molina, H. (1993). Caching and Database Scaling in Distributed Shared-Nothing Information Retrieval Systems. ACM SIGMOD Record, 22(2), 129–138. https://doi.org/10.1145/170036.170063

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free