Data Redaction from Pre-trained GANs
Data Redaction from Pre-trained GANs
复制标题
DOI:
10.1109/satml54575.2023.00048
复制
发表时间:
2022-06
期刊:
影响因子:
--
通讯作者:
Zhifeng Kong;Kamalika Chaudhuri
中科院分区:
文献类型:
--
作者:
Zhifeng Kong;Kamalika Chaudhuri
Large pre-trained generative models are known to occasionally output undesirable samples, which undermines their trustworthiness. The common way to mitigate this is to re-train them differently from scratch using different data or different regularization - which uses a lot of computational resources and does not always fully address the problem. In this work, we take a different, more compute- friendly approach and investigate how to post-edit a model after training so that it “redacts”, or refrains from outputting certain kinds of samples. We show that redaction is a fundamentally different task from data deletion, and data deletion may not always lead to redaction. We then consider Generative Adversar-ial Networks (GANs), and provide three different algorithms for data redaction that differ on how the samples to be redacted are described. Extensive evaluations on real-world image datasets show that our algorithms out-perform data deletion baselines, and are capable of redacting data while retaining high generation quality at a fraction of the cost of full re- training,