Ir al contenido

Documat


A simple method for limiting disclosure in continuous microdata based on principal component analysis

  • Autores: Aída Calviño Martínez
  • Localización: Journal of official statistics, ISSN 0282-423X, Vol. 33, Nº. 1, 2017, págs. 15-41
  • Idioma: inglés
  • DOI: 10.1515/jos-2017-0002
  • Enlaces
  • Resumen
    • In this article we propose a simple and versatile method for limiting disclosure in continuous microdata based on Principal Component Analysis (PCA). Instead of perturbing the original variables, we propose to alter the principal components, as they contain the same information but are uncorrelated, which permits working on each component separately, reducing processing times. The number and weight of the perturbed components determine the level of protection and distortion of the masked data. The method provides preservation of the mean vector and the variance-covariance matrix. Furthermore, depending on the technique chosen to perturb the principal components, the proposed method can provide masked, hybrid or fully synthetic data sets. Some examples of application and comparison with other methods previously proposed in the literature (in terms of disclosure risk and data utility) are also included.


Fundación Dialnet

Mi Documat

Opciones de artículo

Opciones de compartir

Opciones de entorno