RDFSlice: Large-scale RDF Dataset Slicing

  • screenshot

In the last years an increasing number of structured data was published on the Web as Linked Open Data (LOD).Despite recent advances, consuming and using Linked Open Data within an organization is still a substantial challenge. Many of the LOD datasets are quite large and despite progress in RDF data management their loading and querying within a triple store is extremely time-consuming and resource-demanding. To overcome this consumption obstacle, we propose a process inspired by the classical Extract-Transform-Load (ETL) paradigm, RDF dataset slicing.

Download Homepage Source Code

RDFSlicing focuses on the selection and extraction. It devises a fragment of SPARQL dubbed SliceSPARQL, which enables the selection of well-defined slices of datasets fulfilling typical information needs. SliceSPARQL supports graph patterns for which each connected subgraph pattern involves a maximum of one variable or IRI in its join conditions. This restriction guarantees the efficient processing of the query against a sequential dataset dump stream. As a result dataset slices can be generated an order of magnitude faster than by using the conventional approach of loading the whole dataset into a triple store and retrieving the slice by executing the query against the triple store's SPARQL endpoint.

Project Team

Publications

by (Editors: ) [BibTex of ]

News

DBpedia Day @ SEMANTiCS 2022 ( 2022-08-08T11:24:02+02:00 by Julia Holze)

2022-08-08T11:24:02+02:00 by Julia Holze

We are happy to announce that we are partnering again with the SEMANTiCS Conference which will host this year’s DBpedia Day on September 13, 2022. Read more about "DBpedia Day @ SEMANTiCS 2022"

DBpedia Knowledge Engineering PhD Symposium ( 2022-05-02T16:59:37+02:00 by Julia Holze)

2022-05-02T16:59:37+02:00 by Julia Holze

Dear all,  We are excited to invite you to the 1st DBpedia Knowledge Engineering PhD Symposium, organized on July 6th, 2022 in Leipzig, Germany. Read more about "DBpedia Knowledge Engineering PhD Symposium"

Tutorial @ Knowledge Graph Conference 2022 ( 2022-04-25T12:24:06+02:00 by Julia Holze)

2022-04-25T12:24:06+02:00 by Julia Holze

On May 2, 2022 we will organize a tutorial 2.0 at the Knowledge Graph Conference (KGC) 2022. Read more about "Tutorial @ Knowledge Graph Conference 2022"

International Workshop on Data-driven Resilience Research 2022 ( 2022-04-21T14:43:27+02:00 by Julia Holze)

2022-04-21T14:43:27+02:00 by Julia Holze

In the face of continuously changing contextual conditions and ubiquitous disruptive crisis events, the concept of resilience refers to some of the most urgent, challenging, and interesting issues of nowadays society. Read more about "International Workshop on Data-driven Resilience Research 2022"

DBpedia @ Google Summer of Code Program 2022 ( 2022-03-23T14:26:48+01:00 by Julia Holze)

2022-03-23T14:26:48+01:00 by Julia Holze

DBpedia, one of InfAI’s community projects, will be part of the 11th Google Summer of Code (GSoC) program. The GSoC program has the goal to bring students from all over the globe into open source software development. Read more about "DBpedia @ Google Summer of Code Program 2022"