Diversified Stress Testing of RDF Data Management Systems

  • Aluç G
  • Hartig O
  • Özsu M
  • et al.
N/ACitations
Citations of this article
42Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

The Resource Description Framework (RDF) is a standard for con-ceptually describing data on the Web, and SPARQL is the query language for RDF. As RDF data continue to be published across heterogeneous domains and integrated at Web-scale such as in the Linked Open Data (LOD) cloud, RDF data management systems are being exposed to queries that are far more diverse and workloads that are far more varied. The first contribution of our work is an in-depth experimental analysis that shows existing SPARQL benchmarks are not suitable for testing systems for diverse queries and varied workloads. To address these shortcomings, our second contribution is the Waterloo SPARQL Diversity Test Suite (WatDiv) that provides stress testing tools for RDF data management systems. Using WatDiv, we have been able to reveal issues with existing sys-tems that went unnoticed in evaluations using earlier benchmarks. Specifically, our experiments with five popular RDF data management systems show that they cannot deliver good performance uniformly across workloads. For some queries, there can be as much as five orders of magnitude difference between the query execution time of the fastest and the slowest system while the fastest system on one query may unexpectedly time out on another query. By performing a detailed analysis, we pinpoint these problems to specific types of queries and workloads.

Cite

CITATION STYLE

APA

Aluç, G., Hartig, O., Özsu, M. T., & Daudjee, K. (2014). Diversified Stress Testing of RDF Data Management Systems (pp. 197–212). https://doi.org/10.1007/978-3-319-11964-9_13

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free