Skip to Content
Serverless ETL and Analytics with AWS Glue
book

Serverless ETL and Analytics with AWS Glue

by Vishal Pathak, Subramanya Vajiraya, Noritaka Sekiyama, Tomohiro Tanaka, Albert Quiroga, Ishan Gaur
August 2022
Intermediate to advanced
434 pages
10h 34m
English
Packt Publishing
Content preview from Serverless ETL and Analytics with AWS Glue

Chapter 13: Data Analysis

In the previous chapter, we looked at the various buckets of Glue job expectation messages, why they occur, and how to handle them.

We learned about the impact of data skewness, how that can adversely impact job execution, and the techniques you can use to fix it. Additionally, we looked at some of the common reasons for Out-of-Memory (OOM) errors and the out-of-the-box mechanisms that are available in AWS Glue to handle them. Some of these tools and techniques can be used to be more effective in resource utilization in a pay-as-you-go cloud-native world. These techniques can not only be used for efficient processing but also help you reduce the processing time in a world that increasingly needs answers as quickly as ...

Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Start your free trial

You might also like

PySpark and AWS: Master Big Data with PySpark and AWS

PySpark and AWS: Master Big Data with PySpark and AWS

AI Sciences
AWS Certified Data Engineer Associate Study Guide

AWS Certified Data Engineer Associate Study Guide

Sakti Mishra, Dylan Qu, Anusha Challa

Publisher Resources

ISBN: 9781800564985Supplemental Content