From Score Learning to Discretized Sampling: An End-to-End Generalization Analysis of Diffusion Models
Abstract
Despite the empirical success of score-based diffusion models, a complete theoretical understanding of how finite-sample learning, network parameterization, and numerical discretization jointly dictate generative quality remains underdeveloped.
Existing sampling analyses often evaluate the generative performance conditional on an oracle score or a pre-specified error threshold.
In this work, we establish a unified convergence and generalization framework for score-based diffusion models parameterized by practical ResNet-type architectures.
We analyze the generalization and convergence properties from the practical finite-sample, discrete-time learning problem of the score function to the ideal continuous-time, population-level objective.
Based on the generalization result of the learning problem of score function, we analyze the sampling process induced by the learned score function and provide an end-to-end total variation distance estimate for the generated terminal distribution.
This estimate explicitly decomposes the overall generative error into four interpretable components: the truncation error of the forward process, the reverse-time discretization error, the generalization error incorporating both finite data and forward-time discretization, and the training optimization gap.
Our results quantitatively characterize how the training sample size, temporal discretization grids, and optimization accuracy jointly control the final fidelity of samples generated by diffusion models.
이 뉴스, 어떠셨어요?
탭 한 번으로 반응 · 로그인 불필요