Skip to Content
Serverless ETL and Analytics with AWS Glue
book

Serverless ETL and Analytics with AWS Glue

by Vishal Pathak, Subramanya Vajiraya, Noritaka Sekiyama, Tomohiro Tanaka, Albert Quiroga, Ishan Gaur
August 2022
Intermediate to advanced
434 pages
10h 34m
English
Packt Publishing
Content preview from Serverless ETL and Analytics with AWS Glue

Chapter 2: Introduction to Important AWS Glue Features

In the previous chapter, we talked about the evolution of different data management strategies, such as data warehousing, data lakes, the data lakehouse, and data meshes, and the key differences between each. We introduced the Apache Spark framework, briefly discussed the Spark workload execution mechanism, learned how Spark workloads can be fulfilled on the AWS cloud, and introduced AWS Glue and its components.

In this chapter, we will discuss the different components of AWS Glue so that we know how AWS Glue can be used to perform different data integration tasks.

Upon completing this chapter, you will be able to define data integration and explain how AWS Glue can be used for this. You ...

Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Start your free trial

You might also like

PySpark and AWS: Master Big Data with PySpark and AWS

PySpark and AWS: Master Big Data with PySpark and AWS

AI Sciences
AWS Certified Data Engineer Associate Study Guide

AWS Certified Data Engineer Associate Study Guide

Sakti Mishra, Dylan Qu, Anusha Challa

Publisher Resources

ISBN: 9781800564985Supplemental Content