Synthetic datasets for statistical disclosure control : theory and implementation

The aim of this book is to give the reader a detailed introduction to the different approaches to generating multiply imputed synthetic datasets. It describes all approaches that have been developed so far, provides a brief history of synthetic datasets, and gives useful hints on how to deal with re...

Fuld beskrivelse

Enregistré dans:
Bibliografiske detaljer
Hovedforfatter: Drechsler, Jörg
Format: Livre numérique
Sprog:Anglais
Udgivet: New York, NY : Springer New York 2011.
Cham : Springer Nature
Serier:Lecture Notes in Statistics 201
Fag:
Online adgang:Accès sur la plateforme de l'éditeur
Accès sur la plateforme Istex
Accès Université d'Orléans
Accès INSA CVL
Kommentar: Archives Springer e-books (Licence nationale)
Archives Springer e-books (Licence nationale)
Autres localisations: Voir dans le Sudoc
Edition sous un autre format:• Synthetic datasets for statistical disclosure control, theory and implementation, Jörg Drechsler, New York, Springer, 2011, 1 vol. (XX-138 p.), Lecture notes in statistics, 978-1-461-40325-8
• Synthetic Datasets for Statistical Disclosure Control, Texte imprimé, 9781461403272
Beskrivelse
Summary:The aim of this book is to give the reader a detailed introduction to the different approaches to generating multiply imputed synthetic datasets. It describes all approaches that have been developed so far, provides a brief history of synthetic datasets, and gives useful hints on how to deal with real data problems like nonresponse, skip patterns, or logical constraints. Each chapter is dedicated to one approach, first describing the general concept followed by a detailed application to a real dataset providing useful guidelines on how to implement the theory in practice. The discussed multiple imputation approaches include imputation for nonresponse, generating fully synthetic datasets, generating partially synthetic datasets, generating synthetic datasets when the original data is subject to nonresponse, and a two-stage imputation approach that helps to better address the omnipresent trade-off between analytical validity and the risk of disclosure. The book concludes with a glimpse into the future of synthetic datasets, discussing the potential benefits and possible obstacles of the approach and ways to address the concerns of data users and their understandable discomfort with using data that doesn t consist only of the originally collected values.  The book is intended for researchers and practitioners alike. It helps the researcher to find the state of the art in synthetic data summarized in one book with full reference to all relevant papers on the topic. But it is also useful for the practitioner at the statistical agency who is considering the synthetic data approach for data dissemination in the future and wants to get familiar with the topic.
Emne beskrivelse:Archives Springer e-books (Licence nationale)
Archives Springer e-books (Licence nationale)
ISBN:9781461403265
ISSN:2197-7186
Adgang:Accès en ligne pour les établissements français bénéficiaires des licences nationales
Accès soumis à abonnement pour tout autre établissement
Conditions particulières de réutilisation pour les bénéficiaires des licences nationales. chttps://www.licencesnationales.fr/springer-nature-ebooks-contrat-licence-ln-2017