PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 24, 20243 citationsOpen Access

Unlearning Concepts in Diffusion Model via Concept Domain Correction and Concept Preserving Gradient

View Full Paper
YWYongliang WuSZShiji ZhouMYMingzhuo Yang

Key Points

Key points are not available for this paper at this time.

Abstract

Current text-to-image diffusion models have achieved groundbreaking results in image generation tasks. However, the unavoidable inclusion of sensitive information during pre-training introduces significant risks such as copyright infringement and privacy violations in the generated images. Machine Unlearning (MU) provides a effective way to the sensitive concepts captured by the model, has been shown to be a promising approach to addressing these issues. Nonetheless, existing MU methods for concept erasure encounter two primary bottlenecks: 1) generalization issues, where concept erasure is effective only for the data within the unlearn set, and prompts outside the unlearn set often still result in the generation of sensitive concepts; and 2) utility drop, where erasing target concepts significantly degrades the model's performance. To this end, this paper first proposes a concept domain correction framework for unlearning concepts in diffusion models. By aligning the output domains of sensitive concepts and anchor concepts through adversarial training, we enhance the generalizability of the unlearning results. Secondly, we devise a concept-preserving scheme based on gradient surgery. This approach alleviates the parts of the unlearning gradient that contradict the relearning gradient, ensuring that the process of unlearning minimally disrupts the model's performance. Finally, extensive experiments validate the effectiveness of our model, demonstrating our method's capability to address the challenges of concept unlearning in diffusion models while preserving model utility.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wu et al. (2024) studied this question.

synapsesocial.com/papers/68e6886ab6db643587610f3fhttps://doi.org/10.48550/arxiv.2405.15304
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Concept Unlearning by Modeling Key Steps of Diffusion Process2026
  2. 2Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models2024 · 4 citations
  3. 3Probing Unlearned Diffusion Models: A Transferable Adversarial Attack Perspective2024
  4. 4Erasing Concepts from Text-to-Image Diffusion Models with Few-shot Unlearning2024 · 2 citations
  5. 5All but One: Surgical Concept Erasing with Model Preservation in Text-to-Image Diffusion Models2024 · 15 citations