Generalization Dynamics of Linear Diffusion Models

Claudia Merger; Sebastian Goldt

arXiv:2505.24769·stat.ML·February 2, 2026

Generalization Dynamics of Linear Diffusion Models

Claudia Merger, Sebastian Goldt

PDF

TL;DR

This paper analyzes how the hierarchical structure of data affects the generalization of linear diffusion models, revealing regimes where regularization and early stopping improve performance, and quantifying sample complexity effects.

Contribution

It introduces a theoretical framework based on linear neural networks and data covariance spectra to understand diffusion model generalization with finite data.

Findings

01

Hierarchical data structure and regularization mitigate overfitting in low-sample regimes.

02

For large sample sizes, the divergence approaches its optimum linearly with d/N.

03

The analysis clarifies the role of data complexity in diffusion model generalization.

Abstract

Diffusion models are powerful generative models that produce high-quality samples from complex data. While their infinite-data behavior is well understood, their generalization with finite data remains less clear. Classical learning theory predicts that generalization occurs at a sample complexity that is exponential in the dimension, far exceeding practical needs. We address this gap by analyzing diffusion models through the lens of data covariance spectra, which often follow power-law decays, reflecting the hierarchical structure of real data. To understand whether such a hierarchical structure can benefit learning in diffusion models, we develop a theoretical framework based on linear neural networks, congruent with a Gaussian hypothesis on the data. We quantify how the hierarchical organization of variance in the data and regularization impacts generalization. We find two regimes:…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsEarly Stopping · Diffusion