Global Measures of Data Utility for Microdata Masked for Disclosure Limitation

Main Article Content

Mi-Ja Woo
Jerome P. Reiter
https://orcid.org/0000-0002-8374-3832
Anna Oganian
Alan F. Karr

Abstract

When releasing microdata to the public, data disseminators typically alter the original data to protect the confidentiality of database subjects' identities and sensitive attributes. However, such alteration negatively impacts the utility (quality) of the released data. In this paper, we present quantitative measures of data utility for masked microdata, with the aim of improving disseminators' evaluations of competing masking strategies. The measures, which are global in that they reflect similarities between the entire distributions of the original and released data, utilize empirical distribution estimation, cluster analysis, and propensity scores. We evaluate the measures using both simulated and genuine data. The results suggest that measures based on propensity score methods are the most promising for general use.

Article Details

How to Cite
Woo, Mi-Ja, Jerome P. Reiter, Anna Oganian, and Alan F. Karr. 2009. “Global Measures of Data Utility for Microdata Masked for Disclosure Limitation”. Journal of Privacy and Confidentiality 1 (1). https://doi.org/10.29012/jpc.v1i1.568.
Section
Articles

Funding data