A Sharp KL-Convergence Analysis for Diffusion Models under Minimal Assumptions

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Jain, Nishant, Zhang, Tong
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912549095604224
author Jain, Nishant
Zhang, Tong
author_facet Jain, Nishant
Zhang, Tong
contents Diffusion-based generative models have emerged as highly effective methods for synthesizing high-quality samples. Recent works have focused on analyzing the convergence of their generation process with minimal assumptions, either through reverse SDEs or Probability Flow ODEs. The best known guarantees, without any smoothness assumptions, for the KL divergence so far achieve a linear dependence on the data dimension $d$ and an inverse quadratic dependence on $\varepsilon$. In this work, we present a refined analysis that improves the dependence on $\varepsilon$. We model the generation process as a composition of two steps: a reverse ODE step, followed by a smaller noising step along the forward process. This design leverages the fact that the ODE step enables control in Wasserstein-type error, which can then be converted into a KL divergence bound via noise addition, leading to a better dependence on the discretization step size. We further provide a novel analysis to achieve the linear $d$-dependence for the error due to discretizing this Probability Flow ODE in absence of any smoothness assumptions. We show that $\tilde{O}\left(\tfrac{d\log^{3/2}(\frac{1}δ)}{\varepsilon}\right)$ steps suffice to approximate the target distribution corrupted with Gaussian noise of variance $δ$ within $O(\varepsilon^2)$ in KL divergence, improving upon the previous best result, requiring $\tilde{O}\left(\tfrac{d\log^2(\frac{1}δ)}{\varepsilon^2}\right)$ steps.
format Preprint
id arxiv_https___arxiv_org_abs_2508_16306
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle A Sharp KL-Convergence Analysis for Diffusion Models under Minimal Assumptions
Jain, Nishant
Zhang, Tong
Machine Learning
Analysis of PDEs
Statistics Theory
Diffusion-based generative models have emerged as highly effective methods for synthesizing high-quality samples. Recent works have focused on analyzing the convergence of their generation process with minimal assumptions, either through reverse SDEs or Probability Flow ODEs. The best known guarantees, without any smoothness assumptions, for the KL divergence so far achieve a linear dependence on the data dimension $d$ and an inverse quadratic dependence on $\varepsilon$. In this work, we present a refined analysis that improves the dependence on $\varepsilon$. We model the generation process as a composition of two steps: a reverse ODE step, followed by a smaller noising step along the forward process. This design leverages the fact that the ODE step enables control in Wasserstein-type error, which can then be converted into a KL divergence bound via noise addition, leading to a better dependence on the discretization step size. We further provide a novel analysis to achieve the linear $d$-dependence for the error due to discretizing this Probability Flow ODE in absence of any smoothness assumptions. We show that $\tilde{O}\left(\tfrac{d\log^{3/2}(\frac{1}δ)}{\varepsilon}\right)$ steps suffice to approximate the target distribution corrupted with Gaussian noise of variance $δ$ within $O(\varepsilon^2)$ in KL divergence, improving upon the previous best result, requiring $\tilde{O}\left(\tfrac{d\log^2(\frac{1}δ)}{\varepsilon^2}\right)$ steps.
title A Sharp KL-Convergence Analysis for Diffusion Models under Minimal Assumptions
topic Machine Learning
Analysis of PDEs
Statistics Theory
url https://arxiv.org/abs/2508.16306