September 2017
Beginner to intermediate
360 pages
8h 13m
English
Apache Spark is a highly distributed compute engine, which comes with promises of speed and reliability for the computations. As a framework it's based on Hadoop, but it's further enhanced to perform in memory computations to cater to interactive queries and near real-time stream processing. The parallel processing clustering and in-memory processing offer Spark an edge in terms of performance and reliability. Today Apache Spark is known for its proven salient features:
Read now
Unlock full access