Skip to Main Content
CompTIA Data+: DAO-001 Certification Guide
book

CompTIA Data+: DAO-001 Certification Guide

by Cameron Dodd
December 2022
Beginner to intermediate content levelBeginner to intermediate
370 pages
8h 50m
English
Packt Publishing
Content preview from CompTIA Data+: DAO-001 Certification Guide

4

Cleaning and Processing Data

On rare occasions, you may receive data that is already clean, neat, and ready to use, but having an immaculate dataset just handed to you is the exception, not the rule. More often than not, while working as a data analyst, the datasets you receive will be messy, incomplete, and completely unusable without a little work. Trying to use jumbled data will only give you jumbled results. This chapter covers the most common issues you will come across and a few approaches to dealing with them.

Here, we will discuss the difference between duplicate data and redundant data, as well as what to do about it. Then, we will discuss why missing data is an issue and the different approaches you can take to deal with it. Briefly, ...

Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Start your free trial

You might also like

CompTIA Data+: DAO-001 Certification Guide

CompTIA Data+: DAO-001 Certification Guide

Cameron Dodd
CompTIA Data+ DA0-001 Exam Cram

CompTIA Data+ DA0-001 Exam Cram

Akhil Behl, Siva G. Subramanian

Publisher Resources

ISBN: 9781804616086Supplemental Content