Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture introduces the concept of synthetic data generation as a privacy-preserving technique for data publishing. It covers the challenges of anonymization, attribute inference, and identity disclosure in raw datasets. The promise of synthetic data lies in enabling cross-boundary data analytics without compromising privacy. Various generative models and Bayesian networks are discussed, highlighting the importance of protecting customers' sensitive data. The lecture evaluates the privacy gain of publishing synthetic datasets compared to raw datasets, focusing on membership inference and attribute disclosure threats. It concludes that while synthetic data offers some privacy protection, it is not a foolproof solution against privacy threats.
This video is available exclusively on Mediaspace for a restricted audience. Please log in to MediaSpace to access it if you have the necessary permissions.
Watch on Mediaspace