An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Ayed, Fadhel; Battiston, Marco; Camerlenghi, Federico

doi:10.1007/s11222-020-09949-3

Statistics > Methodology

arXiv:2009.11257 (stat)

[Submitted on 23 Sep 2020]

Title:An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Authors:Fadhel Ayed, Marco Battiston, Federico Camerlenghi

View PDF

Abstract:Post Randomization Methods (PRAM) are among the most popular disclosure limitation techniques for both categorical and continuous data. In the categorical case, given a stochastic matrix $M$ and a specified variable, an individual belonging to category $i$ is changed to category $j$ with probability $M_{i,j}$. Every approach to choose the randomization matrix $M$ has to balance between two desiderata: 1) preserving as much statistical information from the raw data as possible; 2) guaranteeing the privacy of individuals in the dataset. This trade-off has generally been shown to be very challenging to solve. In this work, we use recent tools from the computer science literature and propose to choose $M$ as the solution of a constrained maximization problems. Specifically, $M$ is chosen as the solution of a constrained maximization problem, where we maximize the Mutual Information between raw and transformed data, given the constraint that the transformation satisfies the notion of Differential Privacy. For the general Categorical model, it is shown how this maximization problem reduces to a convex linear programming and can be therefore solved with known optimization algorithms.

Subjects:	Methodology (stat.ME); Statistics Theory (math.ST)
Cite as:	arXiv:2009.11257 [stat.ME]
	(or arXiv:2009.11257v1 [stat.ME] for this version)
	https://doi.org/10.48550/arXiv.2009.11257
Related DOI:	https://doi.org/10.1007/s11222-020-09949-3

Submission history

From: Federico Camerlenghi [view email]
[v1] Wed, 23 Sep 2020 17:08:09 UTC (3,516 KB)

Statistics > Methodology

Title:An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Methodology

Title:An Information Theoretic approach to Post Randomization Methods under Differential Privacy

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators