April 2017
Intermediate to advanced
532 pages
12h 39m
English
As discussed in the previous sections, one of the biggest features in the new ML library is the introduction of the pipeline. Pipelines provide a high-level abstraction of the machine learning flow and greatly simplify the complete workflow.
We will demonstrate the process of creating a pipeline in Spark using the StumbleUpon dataset.
Here is a glimpse of the StumbleUpon dataset stored as a temporary table ...
Read now
Unlock full access