Accès libre

A Simple Method for Limiting Disclosure in Continuous Microdata Based on Principal Component Analysis

   | 21 févr. 2017
À propos de cet article

Citez

In this article we propose a simple and versatile method for limiting disclosure in continuous microdata based on Principal Component Analysis (PCA). Instead of perturbing the original variables, we propose to alter the principal components, as they contain the same information but are uncorrelated, which permits working on each component separately, reducing processing times. The number and weight of the perturbed components determine the level of protection and distortion of the masked data. The method provides preservation of the mean vector and the variance-covariance matrix. Furthermore, depending on the technique chosen to perturb the principal components, the proposed method can provide masked, hybrid or fully synthetic data sets. Some examples of application and comparison with other methods previously proposed in the literature (in terms of disclosure risk and data utility) are also included.

eISSN:
2001-7367
Langue:
Anglais
Périodicité:
4 fois par an
Sujets de la revue:
Mathematics, Probability and Statistics