Posts tagged: Synthetic Data

error_bar yellow

How Do We Handle Rare Events in Synthetic Data?

A look into how we can use synthetic data generation to increase the number of rare events in our dataset, and what are risks and benefits of this approach

Read More
histogram orange

Synthetic Data in Machine Learning: Augmentation and Collapse

Exploring the use of synthetic data in machine learning, focusing on augmentation, model collapse, and the implications for health research using synthetic data.

Read More
telescope yellow

How Synthetic Data Gets Made

High level overview of how synthetic data is generated.

Read More
hand blue

Fairness and Bias Amplification in Synthetic Data

Explore how synthetic data can amplify existing biases and affect fairness in health research. Learn why this happens and how it differs from representativeness.

Read More
triangle green

Bias in Synthetic Data

An exploration of bias in synthetic data and its implications for health research.

Read More
shield green

How Private is Synthetic Data? Understanding the Tradeoff with Utility

Synthetic data is a powerful tool for health research, but it comes with a tradeoff between privacy and utility. This blog explores what this means for researchers and how to navigate the tradeoff.

Read More
histogram orange

How Do We Measure the Utility of Synthetic Data?

A practical guide to some of the metrics you can use to evaluate the utility of synthetic data.

Read More
reproducability yellow

Why Synthetic Data is Good for Open Science

Understanding the benefits of synthetic data for open science and reproducibility.

Read More
padlock orange

Is your Synthetic Data actually private?

A practical guide to the three privacy risks in synthetic data, the metrics that quantify them, and why no single number tells you whether your data is safe.

Read More