REVIEW 17 cited by
Denoising Diffusion Bridge Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Diffusion models are powerful generative models that map noise to data using stochastic processes. However, for many applications such as image editing, the model input comes from a distribution that is not random noise. As such, diffusion models must rely on cumbersome methods like guidance or projected sampling to incorporate this information in the generative process. In our work, we propose Denoising Diffusion Bridge Models (DDBMs), a natural alternative to this paradigm based on diffusion bridges, a family of processes that interpolate between two paired distributions given as endpoints. Our method learns the score of the diffusion bridge from data and maps from one endpoint distribution to the other by solving a (stochastic) differential equation based on the learned score. Our method naturally unifies several classes of generative models, such as score-based diffusion models and OT-Flow-Matching, allowing us to adapt existing design and architectural choices to our more general problem. Empirically, we apply DDBMs to challenging image datasets in both pixel and latent space. On standard image translation problems, DDBMs achieve significant improvement over baseline methods, and, when we reduce the problem to image generation by setting the source distribution to random noise, DDBMs achieve comparable FID scores to state-of-the-art methods despite being built for a more general task.
Forward citations
Cited by 17 Pith papers
-
Flowing from Words to Pixels: A Noise-Free Framework for Cross-Modality Evolution
CrossFlow turns text directly into images, and images into text, depth, and higher resolution, by flowing between modality latents without a noise prior or cross-attention.
-
GeoFlow: Efficient Driving Video Generation via Geometry-Aligned Priors
Starting flow-matching video generation from a depth-warped reference frame with spatially-adaptive noise injection reduces required sampling steps by about five times on NuScenes driving videos.
-
AtomDiffuser: Time-Aware Degradation Modeling for Drift and Beam Damage in STEM Imaging
AtomDiffuser predicts affine drift and spatially varying beam damage between STEM image frames, trained on synthetic degradation and shown qualitatively on real cryo-STEM data.
-
TrajDiff: Diffusion Bridge Network with Semantic Alignment for Trajectory Similarity Computation
TrajDiff combines semantic alignment attention, DDBM-based pretraining, and listwise ranking losses to achieve state-of-the-art approximate trajectory similarity on Porto, Geolife, and T-Drive.
-
IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models
A Gaussian-path transition equation lets a pretrained Stable Diffusion model serve as the denoiser inside image restoration bridges, cutting per-task training to a lightweight ControlNet.
-
Training-Free Multi-Step Audio Source Separation
Iteratively remixing and re-separating the input mixture, with the best blend chosen by a quality metric, improves pretrained one-step audio separation models without any retraining.
-
MixBridge: Heterogeneous Image-to-Image Backdoor Attack through Mixture of Schr\"odinger Bridges
MixBridge injects multiple backdoor triggers into image-to-image Schrödinger bridge models by training on poisoned pairs and merging task-specific experts, achieving high attack success and stealthy weights.
-
UniDB: A Unified Diffusion Bridge Framework via Stochastic Optimal Control
A stochastic optimal control formulation of diffusion bridges, where Doob's h-transform is the infinite-penalty limit and a finite penalty yields a tunable detail-preserving bridge.
-
An Ordinary Differential Equation Sampler with Stochastic Start for Diffusion Bridge Models
A stochastic-start ODE sampler for diffusion bridge models avoids the singular start of the probability-flow ODE and beats prior samplers with fewer neural network evaluations.
-
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
OSA-LCM distills a portrait video diffusion model into a single-step generator that matches the quality of a 20-step teacher on FID/FVD, enabling near real-time talking-head generation.
-
Pharmacophore-guided de novo drug design with diffusion bridge
PharmacoBridge, an SE(3)-equivariant diffusion bridge, generates valid 3D drug-like molecules directly from pharmacophore point clouds and outperforms pocket-based baselines in pharmacophore matching and docking affin...
-
Translationese as a Rational Response to Translation Task Difficulty
Translationese is partly predictable from quantifiable translation-task difficulty, especially cross-lingual transfer load, more so for English-to-German than the reverse.
-
Dual guidance: ROM-informed field reconstruction with generative models
Optimized mutual-information sensor placement improves sparse-sensor reconstruction of cylinder-wake flows with a guided diffusion model, but the gains vanish beyond about 25 sensors.
-
Diffusion Bridge or Flow Matching? A Unifying Framework and Comparative Analysis
A theoretical and empirical comparison claiming diffusion bridges have lower stochastic-optimal-control cost and greater robustness than flow matching when training data are scarce.
-
Diffusion Bridge Models for 3D Medical Image Translation
A diffusion bridge model generates 3D T1-to-FA and FA-to-T1 brain images on ADNI data, with downstream classification accuracy close to real images.
-
Force Matching with Relativistic Constraints: A Physics-Inspired Approach to Stable and Efficient Generative Modeling
Force Matching replaces velocity matching in flow-based generative models with a relativistic force objective, but the toy experiments are designed so the model class matches the data generator exactly.
-
Diffusion Prism: Enhancing Diversity and Morphology Consistency in Mask-to-Image Diffusion
Adding small amounts of Gaussian noise and chromatic aberration to binary masks before image-to-image diffusion increases output diversity without losing morphological structure.
Discussion (0). Continue with ORCID to comment.