SPARTex: A Vertex-Centric Framework for RDF Data Analytics

Ibrahim Abdelaziz, Raze Harbi, Semih Salihoglu, Panos Kalnis, Nikos Mamoulis

Research output: Chapter in Book/Report/Conference proceedingConference contribution

16 Scopus citations

Abstract

A growing number of applications require combining SPARQL queries with generic graph search on RDF data. However, the lack of procedural capabilities in SPARQL makes it inappropriate for graph analytics. Moreover, RDF engines focus on SPARQL query evaluation whereas graph management frameworks perform only generic graph computations. In this work, we bridge the gap by introducing SPARTex, an RDF analytics framework based on the vertex-centric computation model. In SPARTex, user-defined vertex centric programs can be invoked from SPARQL as stored procedures. SPARTex allows the execution of a pipeline of graph algorithms without the need for multiple reads/writes of input data and intermediate results. We use a cost-based optimizer for minimizing the communication cost. SPARTex evaluates queries that combine SPARQL and generic graph computations orders of magnitude faster than existing RDF engines. We demonstrate a real system prototype of SPARTex running on a local cluster using real and synthetic datasets. SPARTex has a real-time graphical user interface that allows the participants to write regular SPARQL queries, use our proposed SPARQL extension to declaratively invoke graph algorithms or combine/pipeline both SPARQL querying and generic graph analytics.
Original languageEnglish (US)
Title of host publicationProceedings of the VLDB Endowment
StatePublished - Aug 31 2015

Fingerprint

Dive into the research topics of 'SPARTex: A Vertex-Centric Framework for RDF Data Analytics'. Together they form a unique fingerprint.

Cite this