REVIEW 4 cited by
DreamSampler: Unifying Diffusion Sampling and Score Distillation for Image Manipulation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Reverse sampling and score-distillation have emerged as main workhorses in recent years for image manipulation using latent diffusion models (LDMs). While reverse diffusion sampling often requires adjustments of LDM architecture or feature engineering, score distillation offers a simple yet powerful model-agnostic approach, but it is often prone to mode-collapsing. To address these limitations and leverage the strengths of both approaches, here we introduce a novel framework called {\em DreamSampler}, which seamlessly integrates these two distinct approaches through the lens of regularized latent optimization. Similar to score-distillation, DreamSampler is a model-agnostic approach applicable to any LDM architecture, but it allows both distillation and reverse sampling with additional guidance for image editing and reconstruction. Through experiments involving image editing, SVG reconstruction and etc, we demonstrate the competitive performance of DreamSampler compared to existing approaches, while providing new applications. Code: https://github.com/DreamSampler/dream-sampler
Forward citations
Cited by 4 Pith papers
-
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
A training-free, near-zero-overhead inverse solver for novel-view video generation and inpainting that projects masks into continuous multi-channel latent masks and applies DDS with conjugate gradient in latent space.
-
Inference-Time Diffusion Model Distillation
Distillation++ refines the first denoising steps of distilled diffusion models by interpolating student estimates with teacher model estimates, improving FID and text alignment on several SDXL-based few-step baselines.
-
Optical-Flow Guided Prompt Optimization for Coherent Video Generation
MotionPrompt improves temporal consistency in text-to-video diffusion models by optimizing learnable prompt tokens during sampling, guided by an optical-flow discriminator.
-
FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing
FlowAlign adds a terminal-point source-similarity regularization to inversion-free flow-based editing, improving structural consistency while maintaining semantic alignment.
Discussion (0). Continue with ORCID to comment.