PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 29, 20250 citationsOpen Access

MGD³: Mode-Guided Dataset Distillation using Diffusion Models

View Full Paper
JCJeffrey A. Chan-SantiagoPTPraveen TirupatturGNGaurav Nayak

Key Points

  • The proposed mode-guided method significantly enhances sample diversity and reduces training costs.
  • Accuracy improvements of 4.4% and 2.9% were achieved on benchmark datasets ImageNette and ImageIDC, respectively.
  • Methodology includes three stages: Mode Discovery, Mode Guidance, and Stop Guidance without fine-tuning.
  • Eliminates the necessity for fine-tuning diffusion models, simplifying the distillation process.

Abstract

Dataset distillation has emerged as an effective strategy, significantly reducing training costs and facilitating more efficient model deployment. Recent advances have leveraged generative models to distill datasets by capturing the underlying data distribution. Unfortunately, existing methods require model fine-tuning with distillation losses to encourage diversity and representativeness. However, these methods do not guarantee sample diversity, limiting their performance. We propose a mode-guided diffusion model leveraging a pre-trained diffusion model without the need to fine-tune with distillation losses. Our approach addresses dataset diversity in three stages: Mode Discovery to identify distinct data modes, Mode Guidance to enhance intra-class diversity, and Stop Guidance to mitigate artifacts in synthetic samples that affect performance. Our approach outperforms state-of-the-art methods, achieving accuracy gains of 4.4%, 2.9%, 1.6%, and 1.6% on ImageNette, ImageIDC, ImageNet-100, and ImageNet-1K, respectively. Our method eliminates the need for fine-tuning diffusion models with distillation losses, significantly reducing computational costs. Our code is available on the project webpage: https://jachansantiago.github.io/mode-guided-distillation/

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chan-Santiago et al. (2025) studied this question.

synapsesocial.com/papers/68da58d8c1728099cfd11277https://doi.org/10.48550/arxiv.2505.18963
Ask AI
Helpful
Bookmark
Share
View Full Paper