Sign inSign up

chimerasuite/ontop

By chimerasuite

Updated about 5 years ago

OntopSpark is an Ontop's extension which supports Apache Spark as query processing engine

Image
0

260

chimerasuite/ontop repository overview

Introduction

OntopSpark is an extension of the Ontop Ontology Based Data Access (OBDA) system. OntopSpark uses Apache Spark as a query processing engine for accessing the data stored in data lakes. The integration of a distributed data processing engine such as Apache Spark allows exploiting the Ontop data integration capabilities at its maximum potential, as it brings all the advantages of velocity and parallel computation typical of a distributed system to the task of solving a SPARQL query.

Environment Variables

The OntopSpark docker image has been created starting from the official Ontop source code, therefore the ENV variables to be configured are the same described in the offical Ontop image readme.

Tutorial

For learning how to use OntopSpark inside a data analysis pipeline, you can follow this demo.

Tag summary

Content type

Image

Digest

Size

161 MB

Last updated

about 5 years ago

docker pull chimerasuite/ontop:bfd2333c93