REVIEW 6 cited by
An optimal control perspective on diffusion-based generative modeling
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We establish a connection between stochastic optimal control and generative models based on stochastic differential equations (SDEs), such as recently developed diffusion probabilistic models. In particular, we derive a Hamilton-Jacobi-Bellman equation that governs the evolution of the log-densities of the underlying SDE marginals. This perspective allows to transfer methods from optimal control theory to generative modeling. First, we show that the evidence lower bound is a direct consequence of the well-known verification theorem from control theory. Further, we can formulate diffusion-based generative modeling as a minimization of the Kullback-Leibler divergence between suitable measures in path space. Finally, we develop a novel diffusion-based method for sampling from unnormalized densities -- a problem frequently occurring in statistics and computational sciences. We demonstrate that our time-reversed diffusion sampler (DIS) can outperform other diffusion-based sampling approaches on multiple numerical examples.
Forward citations
Cited by 6 Pith papers
-
From Open Loop to Closed Loop: A Test-Time Iterative Optimization Framework for Reference-Consistent Image Generation
A training-free closed-loop PID controller iteratively corrects latent control signals so diffusion models stay consistent with ID, pose, or depth references better than matched open-loop sampling.
-
Stochastic Quantization as Optimal Control
Stochastic quantization is re-expressed as finite-time optimal control, in which a learned Doob force plus exact path weights reach the Gibbs measure without waiting for equilibrium.
-
FES-FM: Free Energy Surface Sampling via Reduced Flow Matching
FES-FM learns a reduced flow-matching transport in collective-variable space to sample free energy surfaces, cutting per-sample generation cost while leaving full-space training cost unchanged.
-
Solving Inverse Problems with Flow-based Models via Model Predictive Control
MPC-Flow applies model predictive control to guide pretrained flow models through inverse problems, with a single-step variant that avoids backpropagation and scales to 32B-parameter models on consumer hardware.
-
Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning
Diffusion policies can be inserted into maximum-entropy RL by minimizing an upper bound on reverse KL, yielding DiffPPO, DiffSAC, and DiffWPO.
-
Importance Weighted Score Matching for Diffusion Samplers with Enhanced Mode Coverage
Importance Weighted Score Matching trains diffusion samplers by reweighting score matching with self-normalized importance sampling to approximate the forward KL and improve mode coverage.
Discussion (0). Sign in to comment.