An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Home > Research > Publications & Outputs > An Information Theoretic approach to Post Rando...

Mathematics and Statistics

Associated organisational unit

Statistical Artificial Intelligence

Electronic data

S_C_MAIN_FILE
Accepted author manuscript, 7.32 MB, PDF document
Available under license: CC BY: Creative Commons Attribution 4.0 International License

Text available via DOI:

https://doi.org/10.1007/s11222-020-09949-3
Final published version
Available under license: CC BY: Creative Commons Attribution 4.0 International License

Keywords

Post Randomization Methods, Disclosure risk, Mutual Information, Differential Privacy, Categorical Variables

View graph of relations

Research output: Contribution to Journal/Magazine › Journal article › peer-review

Published

Fadhel Ayed
Marco Battiston
Federico Camerlenghi

More...

<mark>Journal publication date</mark>	1/09/2020
<mark>Journal</mark>	Statistics and Computing
Volume	30
Number of pages	15
Pages (from-to)	1347–1361
Publication Status	Published
Early online date	1/06/20
<mark>Original language</mark>	English

Abstract

Post Randomization Methods (PRAM) are among the most popular disclosure limitation techniques for both categorical and continuous data. In the categorical case, given a stochastic matrix M and a specified variable, an individual belonging to category i is changed to category j with probability Mi,j . Every approach to choose the randomization matrix M has to balance between two desiderata: 1) preserving as much statistical information from the raw data as possible; 2) guaranteeing the privacy of individuals in the dataset.
This trade-off has generally been shown to be very challenging to solve. In this work, we use recent tools from the computer science literature and propose to choose M as the solution of a constrained maximization problems. Specifically, M is chosen as the solution of a constrained maximization problem, where we maximize the Mutual Information between raw and transformed data,
given the constraint that the transformation satisfies the notion of Differential Privacy. For the general Categorical model, it is shown how this maximization problem reduces to a convex linear programming and can be therefore solved with known optimization algorithms.

Research

Associated organisational unit

Electronic data

Links

Text available via DOI:

Keywords

An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Abstract

Quick Links

Connect With Us

Faculties & Depts

Contact Us