PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 28, 20240 citationsOpen Access

FineDiffusion: Scaling up Diffusion Models for Fine-grained Image Generation with 10,000 Classes

View Full Paper
ZPZiying PanKWKun WangGLGang Li

Key Points

Key points are not available for this paper at this time.

Abstract

The class-conditional image generation based on diffusion models is renowned for generating high-quality and diverse images. However, most prior efforts focus on generating images for general categories, e.g., 1000 classes in ImageNet-1k. A more challenging task, large-scale fine-grained image generation, remains the boundary to explore. In this work, we present a parameter-efficient strategy, called FineDiffusion, to fine-tune large pre-trained diffusion models scaling to large-scale fine-grained image generation with 10,000 categories. FineDiffusion significantly accelerates training and reduces storage overhead by only fine-tuning tiered class embedder, bias terms, and normalization layers' parameters. To further improve the image generation quality of fine-grained categories, we propose a novel sampling method for fine-grained image generation, which utilizes superclass-conditioned guidance, specifically tailored for fine-grained categories, to replace the conventional classifier-free guidance sampling. Compared to full fine-tuning, FineDiffusion achieves a remarkable 1.56x training speed-up and requires storing merely 1.77% of the total model parameters, while achieving state-of-the-art FID of 9.776 on image generation of 10,000 classes. Extensive qualitative and quantitative experiments demonstrate the superiority of our method compared to other parameter-efficient fine-tuning methods. The code and more generated results are available at our project website: https://finediffusion.github.io/.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Pan et al. (2024) studied this question.

synapsesocial.com/papers/68e7741eb6db6435876e90e8https://doi.org/10.48550/arxiv.2402.18331
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance2024
  2. 2Flash Diffusion: Accelerating Any Conditional Diffusion Model for Few Steps Image Generation2024
  3. 3ReDiFine: Reusable Diffusion Finetuning for Mitigating Degradation in the Chain of Diffusion2024
  4. 4LowDiff: Efficient Diffusion Sampling with Low-Resolution Condition2025
  5. 5Upsample Guidance: Scale Up Diffusion Models without Training2024