REVIEW 3 cited by
Mathematical analysis of singularities in the diffusion model under the submanifold assumption
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper concerns the mathematical analyses of the diffusion model in machine learning. The drift term of the backward sampling process is represented as a conditional expectation involving the data distribution and the forward diffusion. The training process aims to find such a drift function by minimizing the mean-squared residue related to the conditional expectation. Using small-time approximations of the Green's function of the forward diffusion, we show that the analytical mean drift function in DDPM and the score function in SGM asymptotically blow up in the final stages of the sampling process for singular data distributions such as those concentrated on lower-dimensional manifolds, and are therefore difficult to approximate by a network. To overcome this difficulty, we derive a new target function and associated loss, which remains bounded even for singular data distributions. We validate the theoretical findings with several numerical examples.
Forward citations
Cited by 3 Pith papers
-
When and how can inexact generative models still sample from the data manifold?
Inexact generative models stay on the data manifold because infinitesimal learning errors perturb the density only along the manifold, when top Lyapunov vectors align with the support boundary.
-
Memorization and Regularization in Generative Diffusion Models
The exact minimizer of the empirical score-matching loss makes reverse diffusion trajectories converge to training samples, and certain regularizers prevent that collapse.
-
Inconsistencies In Consistency Models: Better ODE Solving Does Not Imply Better Samples
Directly supervising a consistency model against an ODE solver lowers ODE solving error yet degrades image quality, so better ODE solving does not imply better samples.
Discussion (0). Continue with ORCID to comment.