Open sourceActive

Apache Spark

MapReduce-like cluster computing framework for low-latency iterative jobs and interactive use.

Open source page

Field note

What it does

Provides clean, language-integrated APIs in Scala and Java with parallel operators. Runs on Mesos, YARN, EC2, or standalone mode.

Capabilities

Available capabilities

Tags

Tags

Ways to use it

Ways to use it

sdk

Scala / https://spark.apache.org/docs/0.6.1/

sdk

Java / https://spark.apache.org/docs/0.6.1/

sdk

Scala / https://spark.apache.org/docs/0.7.2/

sdk

Python / https://spark.apache.org/docs/0.7.2/

sdk

Scala / https://spark.apache.org/docs/0.6.0/

sdk

Scala / https://spark.apache.org/docs/0.6.2/api/scala/index.html

sdk

Java / https://spark.apache.org/docs/0.6.2/java-programming-guide.html

sdk

Scala / https://spark.apache.org/docs/0.7.3/

sdk

Java / https://spark.apache.org/docs/0.7.3/

sdk

Python / https://spark.apache.org/docs/0.7.3/

Product features

Product features

A Note About Hadoop Versions

Building

Community

Downloading

Heading

Hello, world!

Spark Overview

Testing the Build

Where to Go from Here