Sign inSign up

cambridgesemantics/spark

By cambridgesemantics

•Updated over 4 years ago

Apache spark

Image
0

100K+

cambridgesemantics/spark repository overview

⁠Apache Spark

⁠Important notice

This image provides Apache spark

⁠Supported tags

ReleaseTags
v2.4.8-hadoop2.7-livy0.8.0v2.4.8-hadoop2.7-livy0.8.0, v2.4.8-hadoop2.7-livy0.8.0-r{{ build }}
v2.4.4-hadoop2.7-livy0.8.0v2.4.4-hadoop2.7-livy0.8.0, v2.4.4-hadoop2.7-livy0.8.0-r{{ build }}

⁠About Apache Spark

Spark⁠ is a unified analytics engine for large-scale data processing. It provides high-level APIs in Scala⁠, Java, Python, and R⁠, and an optimized engine that supports general computation graphs for data analysis. It also supports a rich set of higher-level tools including Spark SQL for SQL and DataFrames, MLlib for machine learning, GraphX for graph processing, and Structured Streaming for stream processing.

⁠About Cambridge Semantics⁠

Cambridge Semantics Inc.⁠, The Smart Data Company®, is a big data management and enterprise analytics software company that offers a universal semantic layer to connect and bring meaning to all enterprise data. The company's Anzo® platform for building enterprise data fabrics combines a semantic layer with an embedded graph database to link, analyze and manage diverse data — internal or external, structured or unstructured — at exceptional speed and big data scales. AnzoGraph, the same graph database embedded within Anzo, is now available on a standalone basis for data analysts, enterprise architects, and application developers to build and execute data warehouse analytics, graph & data science algorithms, and inferencing, all in one award-winning graph database for analytics.

Tag summary

Content type

Image

Digest

Size

354.4 MB

Last updated

over 4 years ago

docker pull cambridgesemantics/spark