Back to Search Start Over

Diversified Stress Testing of RDF Data Management Systems

Authors :
Olaf Hartig
Güneş Aluç
Khuzaima Daudjee
M. Tamer Özsu
Source :
The Semantic Web – ISWC 2014 ISBN: 9783319119632, Semantic Web Conference (1)
Publication Year :
2014
Publisher :
Springer International Publishing, 2014.

Abstract

The Resource Description Framework (RDF) is a standard for conceptually describing data on the Web, and SPARQL is the query language for RDF. As RDF data continue to be published across heterogeneous domains and integrated at Web-scale such as in the Linked Open Data (LOD) cloud, RDF data management systems are being exposed to queries that are far more diverse and workloads that are far more varied. The first contribution of our work is an in-depth experimental analysis that shows existing SPARQL benchmarks are not suitable for testing systems for diverse queries and varied workloads. To address these shortcomings, our second contribution is the Waterloo SPARQL Diversity Test Suite (WatDiv) that provides stress testing tools for RDF data management systems. Using WatDiv, we have been able to reveal issues with existing systems that went unnoticed in evaluations using earlier benchmarks. Specifically, our experiments with five popular RDF data management systems show that they cannot deliver good performance uniformly across workloads. For some queries, there can be as much as five orders of magnitude difference between the query execution time of the fastest and the slowest system while the fastest system on one query may unexpectedly time out on another query. By performing a detailed analysis, we pinpoint these problems to specific types of queries and workloads.

Details

ISBN :
978-3-319-11963-2
ISBNs :
9783319119632
Database :
OpenAIRE
Journal :
The Semantic Web – ISWC 2014 ISBN: 9783319119632, Semantic Web Conference (1)
Accession number :
edsair.doi...........69857e4a5d3f2286ad68a39f7288ddde
Full Text :
https://doi.org/10.1007/978-3-319-11964-9_13