PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 8, 20240 citationsOpen Access

Imagine Flash: Accelerating Emu Diffusion Models with Backward Distillation

View Full Paper
JKJonas KöhlerAPAlbert PumarolaESEdgar Schönfeld

Key Points

Key points are not available for this paper at this time.

Abstract

Diffusion models are a powerful generative framework, but come with expensive inference. Existing acceleration methods often compromise image quality or fail under complex conditioning when operating in an extremely low-step regime. In this work, we propose a novel distillation framework tailored to enable high-fidelity, diverse sample generation using just one to three steps. Our approach comprises three key components: (i) Backward Distillation, which mitigates training-inference discrepancies by calibrating the student on its own backward trajectory; (ii) Shifted Reconstruction Loss that dynamically adapts knowledge transfer based on the current time step; and (iii) Noise Correction, an inference-time technique that enhances sample quality by addressing singularities in noise prediction. Through extensive experiments, we demonstrate that our method outperforms existing competitors in quantitative metrics and human evaluations. Remarkably, it achieves performance comparable to the teacher model using only three denoising steps, enabling efficient high-quality generation.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Köhler et al. (2024) studied this question.

synapsesocial.com/papers/68e6b14fb6db6435876333c0https://doi.org/10.48550/arxiv.2405.05224
Ask AI
Helpful
Bookmark
Share
View Full Paper