Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Dec 09, 2024

Myeongseob Ko, Henry Li, Zhun Wang, Jonathan Patsenker, Jiachen T. Wang, Qinbin Li, Ming Jin, Dawn Song, Ruoxi Jia

Figure 1 for Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Figure 2 for Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Figure 3 for Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Figure 4 for Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Share this with someone who'll enjoy it:

Abstract:Large-scale generative models have shown impressive image-generation capabilities, propelled by massive data. However, this often inadvertently leads to the generation of harmful or inappropriate content and raises copyright concerns. Driven by these concerns, machine unlearning has become crucial to effectively purge undesirable knowledge from models. While existing literature has studied various unlearning techniques, these often suffer from either poor unlearning quality or degradation in text-image alignment after unlearning, due to the competitive nature of these objectives. To address these challenges, we propose a framework that seeks an optimal model update at each unlearning iteration, ensuring monotonic improvement on both objectives. We further derive the characterization of such an update. In addition, we design procedures to strategically diversify the unlearning and remaining datasets to boost performance improvement. Our evaluation demonstrates that our method effectively removes target classes from recent diffusion-based generative models and concepts from stable diffusion models while maintaining close alignment with the models' original trained states, thus outperforming state-of-the-art baselines. Our code will be made available at \url{https://github.com/reds-lab/Restricted_gradient_diversity_unlearning.git}.

* 22 pages, The Thirty-Eighth Annual Conference on Neural Information Processing Systems

View paper on

Share this with someone who'll enjoy it:

Title:Boosting Alignment for Post-Unlearning Text-to-Image Generative Models

Paper and Code