Skip to Content
An Introduction to Machine Learning Models in Production
on-demand course

An Introduction to Machine Learning Models in Production

with Jason Slepicka
December 2017
Intermediate
39m
English
O'Reilly Media, Inc.
Closed Captioning available in German, English, Spanish, French, Japanese, Korean, Portuguese (Portugal, Brazil), Chinese (Simplified), Chinese (Traditional)

Overview

This course lays out the common architecture, infrastructure, and theoretical considerations for managing an enterprise machine learning (ML) model pipeline. Because automation is the key to effective operations, you'll learn about open source tools like Spark, Hive, ModelDB, and Docker and how they're used to bridge the gap between individual models and a reproducible pipeline. You'll also learn how effective data teams operate; why they use a common process for building, training, deploying, and maintaining ML models; and how they're able to seamlessly push models into production. The course is designed for the data engineer transitioning to the cloud and for the data scientist ready to use model deployment pipelines that are reproducible and automated. Learners should have basic familiarity with: cloud platforms like Amazon Web Services; Scala or Python; Hadoop, Spark, or Pandas; SBT or Maven; Bash, Docker, and REST.

  • Understand how to set-up and manage an enterprise ML model pipeline
  • Learn the common components that make up enterprise ML model pipelines
  • Explore the use and purpose of pipeline tools like Spark, Hive, ModelDB, and Docker
  • Discover the gaps in the Spark ecosystem for maintaining and deploying ML pipelines
  • Learn how to move from creating one-off models to building a reproducible automated pipeline

Jason Slepicka is a senior data engineer with Los Angeles based DataScience, where he builds pipelines and data science platform infrastructure. He has a decade of experience integrating data to support efforts like fighting human trafficking for DARPA, exploring the evolution of evolvability in yeast, and tracking intruders in computer networks. Jason has both a Bachelor's and Master’s in Computer Science from the University of Arizona and is working on his PhD in Computer Science at the University of Southern California Information Sciences Institute.

Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.

Watch now

Unlock full access

More than 5,000 organizations count on O’Reilly

AirBnbBlueOriginElectronic ArtsHomeDepotNasdaqRakutenTata Consultancy Services

QuotationMarkO’Reilly covers everything we've got, with content to help us build a world-class technology community, upgrade the capabilities and competencies of our teams, and improve overall team performance as well as their engagement.
Julian F.
Head of Cybersecurity
QuotationMarkI wanted to learn C and C++, but it didn't click for me until I picked up an O'Reilly book. When I went on the O’Reilly platform, I was astonished to find all the books there, plus live events and sandboxes so you could play around with the technology.
Addison B.
Field Engineer
QuotationMarkI’ve been on the O’Reilly platform for more than eight years. I use a couple of learning platforms, but I'm on O'Reilly more than anybody else. When you're there, you start learning. I'm never disappointed.
Amir M.
Data Platform Tech Lead
QuotationMarkI'm always learning. So when I got on to O'Reilly, I was like a kid in a candy store. There are playlists. There are answers. There's on-demand training. It's worth its weight in gold, in terms of what it allows me to do.
Mark W.
Embedded Software Engineer

You might also like

AI Superstream: Efficient Machine Learning

AI Superstream: Efficient Machine Learning

Shingai Manjengwa, Jesse Hoey, Michael Houston, Xin Li, Maryam Mehri Dehnavi, Meena Arunachalam, Moty Fania, Alishba Imran

Publisher Resources

ISBN: 9781491988794