OntopSpark is an Ontop's extension which supports Apache Spark as query processing engine
260
OntopSpark is an extension of the Ontop Ontology Based Data Access (OBDA) system. OntopSpark uses Apache Spark as a query processing engine for accessing the data stored in data lakes. The integration of a distributed data processing engine such as Apache Spark allows exploiting the Ontop data integration capabilities at its maximum potential, as it brings all the advantages of velocity and parallel computation typical of a distributed system to the task of solving a SPARQL query.
The OntopSpark docker image has been created starting from the official Ontop source code, therefore the ENV variables to be configured are the same described in the offical Ontop image readme.
For learning how to use OntopSpark inside a data analysis pipeline, you can follow this demo.
Content type
Image
Digest
Size
161 MB
Last updated
about 5 years ago
docker pull chimerasuite/ontop:bfd2333c93