This image provides Apache spark
| Release | Tags |
|---|---|
| v2.4.8-hadoop2.7-livy0.8.0 | v2.4.8-hadoop2.7-livy0.8.0, v2.4.8-hadoop2.7-livy0.8.0-r{{ build }} |
| v2.4.4-hadoop2.7-livy0.8.0 | v2.4.4-hadoop2.7-livy0.8.0, v2.4.4-hadoop2.7-livy0.8.0-r{{ build }} |
Spark is a unified analytics engine for large-scale data processing. It provides high-level APIs in Scala, Java, Python, and R, and an optimized engine that supports general computation graphs for data analysis. It also supports a rich set of higher-level tools including Spark SQL for SQL and DataFrames, MLlib for machine learning, GraphX for graph processing, and Structured Streaming for stream processing.
Cambridge Semantics Inc., The Smart Data Company®, is a big data management and enterprise analytics software company that offers a universal semantic layer to connect and bring meaning to all enterprise data. The company's Anzo® platform for building enterprise data fabrics combines a semantic layer with an embedded graph database to link, analyze and manage diverse data — internal or external, structured or unstructured — at exceptional speed and big data scales. AnzoGraph, the same graph database embedded within Anzo, is now available on a standalone basis for data analysts, enterprise architects, and application developers to build and execute data warehouse analytics, graph & data science algorithms, and inferencing, all in one award-winning graph database for analytics.
Content type
Image
Digest
Size
354.4 MB
Last updated
over 4 years ago
docker pull cambridgesemantics/spark