CSVImport: Representing multi-dimensional statistical data as RDF using the RDF Data Cube Vocabulary

  • screenshot

This project is about the representation of multi-dimensional statistical data as RDF using the RDF Data Cube vocabulary by importing spreadsheets into the OntoWiki plugin.

Homepage Issues Wiki

Statistical data on the web is often published as Excel sheets. Although they have the advantage of being easily readable by humans, they cannot be queried efficiently. Also it is difficult to integrate with other datasets, which may be in different formats. Our approach is to convert the data into a single data model – RDF. But in these datasets, a single statistical value is described in several dimensions. Thus a simple row-based transformation is not possible. Therefore, we use The RDF Data Cube vocabulary for the conversion as it is designed particularly to represent multidimensional statistical data using RDF. Transforming CSV to RDF in a fully automated way is not feasible as there may be dimensions encoded in the heading or label of a sheet. Therefore, we introduce a semi-automated approach as a plugin in OntoWiki.

Project Team

Former Members

News

DBpedia @ Google Summer of Code – GSoC 2017 ( 2017-03-13T11:12:50+01:00 Christopher Schulz)

2017-03-13T11:12:50+01:00 Christopher Schulz

DBpedia, one of InfAI’s community projects, will be part of the 5th Google Summer of Code program. The GsoC has the goal to bring students from all over the globe into open source software development. Read more about "DBpedia @ Google Summer of Code – GSoC 2017"

New GERBIL release v1.2.5 – Benchmarking entity annotation systems ( 2017-03-10T11:49:51+01:00 by Ricardo Usbeck)

2017-03-10T11:49:51+01:00 by Ricardo Usbeck

Dear all, the Smart Data Management competence center at AKSW is happy to announce GERBIL 1.2.5. Read more about "New GERBIL release v1.2.5 – Benchmarking entity annotation systems"

DBpedia Open Text Extraction Challenge – TextExt ( 2017-03-09T12:15:57+01:00 Christopher Schulz)

2017-03-09T12:15:57+01:00 Christopher Schulz

DBpedia, a community project affiliated with the Institute for Applied Informatics (InfAI) e.V., extract structured information from Wikipedia & Wikidata. Now DBpedia started the DBpedia Open Text Extraction Challenge – TextExt. Read more about "DBpedia Open Text Extraction Challenge – TextExt"

The USPTO Linked Patent Dataset release ( 2017-02-24T17:18:51+01:00 by Mofeed Hassan)

2017-02-24T17:18:51+01:00 by Mofeed Hassan

Dear all, We are happy to announce USPTO Linked Patent Dataset release. Patents are widely used to protect intellectual property and a measure of innovation output. Read more about "The USPTO Linked Patent Dataset release"

Two accepted papers in ESWC 2017 ( 2017-02-22T17:43:38+01:00 by Dr. Mohamed Ahmed Sherif)

2017-02-22T17:43:38+01:00 by Dr. Mohamed Ahmed Sherif

Hello Community! We are very pleased to announce the acceptance of two papers in ESWC 2017 research track. The ESWC 2017 is to be held in Portoroz, Slovenia from 28th of May to the 1st of June. Read more about "Two accepted papers in ESWC 2017"