December 2023
Intermediate to advanced
310 pages
7h 34m
English
In the previous chapter, we talked about how to build a data governance program for our organization and how to identify types of sensitive data. Our work does not stop there. Although in some cases we can safely exclude sensitive information, other times we cannot. So, our machine learning (ML) models that solve problems might need to contain personal data. Sometimes that data can be relevant and useful, or it can create unintended correlations that make the model biased. This is the issue that we will tackle in this chapter.
We will talk about how to recognize sensitive information and how to mitigate it if it is not relevant to the model training process by using techniques such as differential ...
Read now
Unlock full access