Continual Release of Differentially Private Synthetic Data from
  Longitudinal Data Collections

Bun, Mark; Gaboardi, Marco; Neunhoeffer, Marcel; Zhang, Wanrong

Continual Release of Differentially Private Synthetic Data from Longitudinal Data Collections

Authors: Mark Bun
Marco Gaboardi
Marcel Neunhoeffer
Wanrong Zhang
Publication date: 14 May 2024
Publisher
Doi

Abstract

Motivated by privacy concerns in long-term longitudinal studies in medical and social science research, we study the problem of continually releasing differentially private synthetic data from longitudinal data collections. We introduce a model where, in every time step, each individual reports a new data element, and the goal of the synthesizer is to incrementally update a synthetic dataset in a consistent way to capture a rich class of statistical properties. We give continual synthetic data generation algorithms that preserve two basic types of queries: fixed time window queries and cumulative time queries. We show nearly tight upper bounds on the error rates of these algorithms and demonstrate their empirical performance on realistically sized datasets from the U.S. Census Bureau's Survey of Income and Program Participation

Similar works

Full text

Available Versions

Boston University Institutional Repository (OpenBU)

oai:open.bu.edu:2144/50291

Last time updated on 19/07/2025

arXiv.org e-Print Archive

oai:arXiv.org:2306.07884

Last time updated on 15/06/2023