Efficient Diffusion Training via Min-SNR Weighting Strategy

Tiankai Hang; Shuyang Gu; Chen Li; Jianmin Bao; Dong Chen; Han Hu; Xin; Geng; Baining Guo

arXiv:2303.09556·cs.CV·March 12, 2024·5 cites

Efficient Diffusion Training via Min-SNR Weighting Strategy

Tiankai Hang, Shuyang Gu, Chen Li, Jianmin Bao, Dong Chen, Han Hu, Xin, Geng, Baining Guo

PDF

Open Access 2 Repos 6 Models 1 Video

TL;DR

This paper introduces Min-SNR-$\gamma$ weighting for diffusion model training, significantly accelerating convergence and improving image generation quality on ImageNet benchmarks.

Contribution

It proposes a novel Min-SNR-$\gamma$ weighting strategy that balances timestep conflicts, leading to faster and more effective diffusion model training.

Findings

01

3.4× faster convergence than previous methods

02

Achieved a new FID score of 2.06 on ImageNet 256×256

03

More effective with smaller architectures

Abstract

Denoising diffusion models have been a mainstream approach for image generation, however, training these models often suffers from slow convergence. In this paper, we discovered that the slow convergence is partly due to conflicting optimization directions between timesteps. To address this issue, we treat the diffusion training as a multi-task learning problem, and introduce a simple yet effective approach referred to as Min-SNR- $γ$ . This method adapts loss weights of timesteps based on clamped signal-to-noise ratios, which effectively balances the conflicts among timesteps. Our results demonstrate a significant improvement in converging speed, 3.4 $\times$ faster than previous weighting strategies. It is also more effective, achieving a new record FID score of 2.06 on the ImageNet $256 \times 256$ benchmark using smaller architectures than that employed in previous state-of-the-art.…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Models

Videos

Efficient Diffusion Training via Min-SNR Weighting Strategy· youtube

Taxonomy

TopicsDomain Adaptation and Few-Shot Learning · Generative Adversarial Networks and Image Synthesis · Advanced Neuroimaging Techniques and Applications

MethodsDiffusion