Database Anonymization

Book description

The current social and economic context increasingly demands open data to improve scientific research and decision making. However, when published data refer to individual respondents, disclosure risk limitation techniques must be implemented to anonymize the data and guarantee by design the fundamental right to privacy of the subjects the data refer to. Disclosure risk limitation has a long record in the statistical and computer science research communities, who have developed a variety of privacy-preserving solutions for data releases. This Synthesis Lecture provides a comprehensive overview of the fundamentals of privacy in data releases focusing on the computer science perspective. Specifically, we detail the privacy models, anonymization methods, and utility and risk metrics that have been proposed so far in the literature. Besides, as a more advanced topic, we identify and discuss in detail connections between several privacy models (i.e., how to accumulate the privacy guarantees they offer to achieve more robust protection and when such guarantees are equivalent or complementary); we also explore the links between anonymization methods and privacy models (how anonymization methods can be used to enforce privacy models and thereby offer ex ante privacy guarantees). These latter topics are relevant to researchers and advanced practitioners, who will gain a deeper understanding on the available data anonymization solutions and the privacy guarantees they can offer.

Table of contents

  1. Preface
  2. Acknowledgments
  3. Introduction
  4. Privacy in Data Releases
    1. Types of Data Releases
    2. Microdata Sets
    3. Formalizing Privacy
    4. Disclosure Risk in Microdata Sets
    5. Microdata Anonymization
    6. Measuring Information Loss
    7. Trading Off Information Loss and Disclosure Risk
    8. Summary
  5. Anonymization Methods for Microdata
    1. Non-perturbative Masking Methods
    2. Perturbative Masking Methods
    3. Synthetic Data Generation
    4. Summary
  6. Quantifying Disclosure Risk: Record Linkage
    1. Threshold-based Record Linkage
    2. Rule-based Record Linkage
    3. Probabilistic Record Linkage
    4. Summary
  7. The k-Anonymity Privacy Model
    1. Insufficiency of Data De-identification
    2. The k-Anonymity Model
    3. Generalization and Suppression Based k-Anonymity (1/2)
    4. Generalization and Suppression Based k-Anonymity (2/2)
    5. Microaggregation-based k-Anonymity
    6. Probabilistic k-Anonymity
    7. Summary
  8. Beyond k-Anonymity: l-Diversity and t-Closeness
    1. l-Diversity
    2. t-Closeness
    3. Summary
  9. t-Closeness Through Microaggregation
    1. Standard Microaggregation and Merging
    2. t-Closeness Aware Microaggregation: k-anonymity-first
    3. t-Closeness Aware Microaggregation: t-closeness-first
    4. Summary
  10. Differential Privacy
    1. Definition
    2. Calibration to the Global Sensitivity
    3. Calibration to the Smooth Sensitivity
    4. The Exponential Mechanism
    5. Relation to k-anonymity-based Models
    6. Differentially Private Data Publishing
    7. Summary
  11. Differential Privacy by Multivariate Microaggregation
    1. Reducing Sensitivity Via Prior Multivariate Microaggregation
    2. Differentially Private Data Sets by Insensitive Microaggregation
    3. General Insensitive Microaggregation
    4. Differential Privacy with Categorical Attributes
    5. A Semantic Distance for Differential Privacy
    6. Integrating Heterogeneous Attribute Types
    7. Summary
  12. Differential Privacy by Individual Ranking Microaggregation
    1. Limitations of Multivariate Microaggregation
    2. Sensitivity Reduction Via Individual Ranking
    3. Choosing the Microggregation Parameter k
    4. Summary
  13. Conclusions and Research Directions
    1. Summary and Conclusions
    2. Research Directions
  14. Bibliography (1/2)
  15. Bibliography (2/2)
  16. Authors' Biographies
  17. Blank Page (1/3)
  18. Blank Page (2/3)
  19. Blank Page (3/3)

Product information

  • Title: Database Anonymization
  • Author(s): Josep Domingo-Ferrer, David Sánchez, Jordi Soria-Comas
  • Release date: January 2016
  • Publisher(s): Morgan & Claypool Publishers
  • ISBN: 9781627058445