Sciweavers

SIGMOD
2006
ACM

Injecting utility into anonymized datasets

14 years 11 months ago
Injecting utility into anonymized datasets
Limiting disclosure in data publishing requires a careful balance between privacy and utility. Information about individuals must not be revealed, but a dataset should still be useful for studying the characteristics of a population. Privacy requirements such as k-anonymity and -diversity are designed to thwart attacks that attempt to identify individuals in the data and to discover their sensitive information. On the other hand, the utility of such data has not been well-studied. In this paper we will discuss the shortcomings of current heuristic approaches to measuring utility and we will introduce a formal approach to measuring utility. Armed with this utility metric, we will show how to inject additional information into k-anonymous and -diverse tables. This information has an intuitive semantic meaning, it increases the utility beyond what is possible in the original k-anonymity and -diversity frameworks, and it maintains the privacy guarantees of k-anonymity and -diversity.
Daniel Kifer, Johannes Gehrke
Added 08 Dec 2009
Updated 08 Dec 2009
Type Conference
Year 2006
Where SIGMOD
Authors Daniel Kifer, Johannes Gehrke
Comments (0)