MARATTO

article · Frontiers in Digital Health

Synthetic data generation: a privacy-preserving approach to accelerate rare disease research

202528 citationsOpen accessNile University

Abstract

Rare disease research faces significant challenges due to limited patient data, strict privacy regulations, and the need for diverse datasets to develop accurate AI-driven diagnostics and treatments. Synthetic data-artificially generated datasets that mimic patient data while preserving privacy-offer a promising solution to these issues. This article explores how synthetic data can bridge data gaps, enabling the training of AI models, simulating clinical trials, and facilitating cross-border collaborations in rare disease research. We examine case studies where synthetic data successfully replicated patient characteristics, and supported predictive modelling and ensured compliance with regulations like GDPR and HIPAA. While acknowledging current limitations, we discuss synthetic data's potential to revolutionise rare disease research by enhancing data availability and privacy file enabling more efficient and effective research efforts in diagnosing, treating, and managing rare diseases globally.

Research topics

  • Privacy-Preserving Technologies in Data
  • Ethics in Clinical Research
  • Artificial Intelligence in Healthcare and Education

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.3389/fdgth.2025.1563991

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.