Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access

Mudit Gaur; Prashant Trivedi; Sasidhar Kunapuli; Amrit Singh Bedi; and Vaneet Aggarwal

arXiv:2505.18344·cs.LG·April 14, 2026

Improved Sample Complexity For Diffusion Model Training Without Empirical Risk Minimizer Access

Mudit Gaur, Prashant Trivedi, Sasidhar Kunapuli, Amrit Singh Bedi, and Vaneet Aggarwal

PDF

TL;DR

This paper offers a new theoretical analysis of diffusion models, establishing a sample complexity bound of O(ε^{-4}) without requiring access to empirical risk minimizers, improving understanding of their training efficiency.

Contribution

It provides the first sample complexity bound for diffusion model training that does not assume access to the empirical risk minimizer, using a structured error decomposition.

Findings

01

Achieves a sample complexity of O(ε^{-4}) for score estimation.

02

Eliminates exponential dependence on neural network parameters in analysis.

03

First result to remove the need for empirical risk minimizer access in this context.

Abstract

Diffusion models have demonstrated state-of-the-art performance across vision, language, and scientific domains. Despite their empirical success, prior theoretical analyses of the sample complexity suffer from poor scaling with input data dimension or rely on unrealistic assumptions such as access to exact empirical risk minimizers. In this work, we provide a principled analysis of score estimation, establishing a sample complexity bound of $O (ϵ^{- 4})$ . Our approach leverages a structured decomposition of the score estimation error into statistical, approximation, and optimization errors, enabling us to eliminate the exponential dependence on neural network parameters that arises in prior analyses. It is the first such result that achieves sample complexity bounds without assuming access to the empirical risk minimizer of score function estimation loss.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.