July 2017
Intermediate to advanced
796 pages
18h 55m
English
To provide you with more advanced and additional big data processing capabilities, your Spark jobs can be running on top of Hadoop-based (aka YARN) or Mesos-based clusters. On the other hand, the core APIs in Spark, which is written in Scala, enable you to develop your Spark application using several programming languages such as Java, Scala, Python, and R. Spark provides several libraries that are part of the Spark ecosystems for additional capabilities for general purpose data processing and analytics, graph processing, large-scale structured SQL, and machine learning (ML) areas. The Spark ecosystem consists of the following components:
The core engine of Spark is written ...
Read now
Unlock full access