JOURNAL ARTICLE

Differentially Private Synthetic Data Generation Using Context-Aware GANs

Abstract

The widespread use of big data across sectors has raised major privacy concerns, especially when sensitive information is shared or analyzed. Regulations such as GDPR and HIPAA impose strict controls on data handling, making it difficult to balance the need for insights with privacy requirements. Synthetic data offers a promising solution by creating artificial datasets that reflect real patterns without exposing sensitive information. However, traditional synthetic data methods often fail to capture complex, implicit rules that link different elements of the data and are essential in domains like healthcare. They may reproduce explicit patterns but overlook domain-specific constraints that are not directly stated yet crucial for realism and utility. For example, prescription guidelines that restrict certain medications for specific conditions or prevent harmful drug interactions may not appear explicitly in the original data. Synthetic data generated without these implicit rules can lead to medically inappropriate or unrealistic profiles. To address this gap, we propose ContextGAN, a Context-Aware Differentially Private Generative Adversarial Network that integrates domain-specific rules through a constraint matrix encoding both explicit and implicit knowledge. The constraint-aware discriminator evaluates synthetic data against these rules to ensure adherence to domain constraints, while differential privacy protects sensitive details from the original data. We validate ContextGAN across healthcare, security, and finance, showing that it produces high-quality synthetic data that respects domain rules and preserves privacy. Our results demonstrate that ContextGAN improves realism and utility by enforcing domain constraints, making it suitable for applications that require compliance with both explicit patterns and implicit rules under strict privacy guarantees.

Keywords:
Computer science Context (archaeology)

Metrics

0
Cited By
0.00
FWCI (Field Weighted Citation Impact)
47
Refs
0.25
Citation Normalized Percentile
Is in top 1%
Is in top 10%

Topics

Privacy-Preserving Technologies in Data
Physical Sciences →  Computer Science →  Artificial Intelligence
Cryptography and Data Security
Physical Sciences →  Computer Science →  Artificial Intelligence
Advanced Data Storage Technologies
Physical Sciences →  Computer Science →  Computer Networks and Communications

Related Documents

JOURNAL ARTICLE

Differentially private synthetic medical data generation using convolutional GANs

Amirsina TorfiEdward A. FoxChandan K. Reddy

Journal:   Information Sciences Year: 2021 Vol: 586 Pages: 485-500
JOURNAL ARTICLE

Differentially private GANs for generating synthetic indoor location data

Vahideh MoghtadaieeMina AlishahiMilad Rabiei

Journal:   International Journal of Information Security Year: 2025 Vol: 24 (3)
BOOK-CHAPTER

TraVaG: Differentially Private Trace Variant Generation Using GANs

Majid RafieiFrederik WangelikMahsa PourbafraniWil M. P. van der Aalst

Lecture notes in business information processing Year: 2023 Pages: 415-431
JOURNAL ARTICLE

Online Differentially Private Synthetic Data Generation

Yiyun HeRoman VershyninYizhe Zhu

Journal:   IEEE Transactions on Privacy Year: 2024 Vol: 1 Pages: 19-30
© 2026 ScienceGate Book Chapters — All rights reserved.