REVIEW 4 cited by
SinSR: Diffusion-Based Image Super-Resolution in a Single Step
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
While super-resolution (SR) methods based on diffusion models exhibit promising results, their practical application is hindered by the substantial number of required inference steps. Recent methods utilize degraded images in the initial state, thereby shortening the Markov chain. Nevertheless, these solutions either rely on a precise formulation of the degradation process or still necessitate a relatively lengthy generation path (e.g., 15 iterations). To enhance inference speed, we propose a simple yet effective method for achieving single-step SR generation, named SinSR. Specifically, we first derive a deterministic sampling process from the most recent state-of-the-art (SOTA) method for accelerating diffusion-based SR. This allows the mapping between the input random noise and the generated high-resolution image to be obtained in a reduced and acceptable number of inference steps during training. We show that this deterministic mapping can be distilled into a student model that performs SR within only one inference step. Additionally, we propose a novel consistency-preserving loss to simultaneously leverage the ground-truth image during the distillation process, ensuring that the performance of the student model is not solely bound by the feature manifold of the teacher model, resulting in further performance improvement. Extensive experiments conducted on synthetic and real-world datasets demonstrate that the proposed method can achieve comparable or even superior performance compared to both previous SOTA methods and the teacher model, in just one sampling step, resulting in a remarkable up to x10 speedup for inference. Our code will be released at https://github.com/wyf0912/SinSR
Forward citations
Cited by 4 Pith papers
-
When Latents Forget Pixels: Restoring Fidelity in Diffusion Transformer Super-Resolution
By injecting pre-VAE pixel features into both the latent denoising trajectory and the frozen VAE decoder, PGSR improves fidelity of latent diffusion transformer super-resolution while maintaining perceptual quality.
-
GuideSR: Rethinking Guidance for One-Step High-Fidelity Diffusion-Based Super-Resolution
GuideSR pairs a full-resolution guidance branch with a one-step latent diffusion branch and improves PSNR, SSIM, LPIPS, DISTS, and FID on DIV2K-Val, RealSR, and DRealSR.
-
D$^2$-DPM: Dual Denoising for Quantized Diffusion Probabilistic Models
Modeling quantization noise in compressed diffusion models as a time-step-dependent joint Gaussian, then correcting its mean and variance during sampling, improves FID over prior PTQ methods and can beat the full-prec...
-
Diffusion Model Quantization: A Review
A structured review and benchmark of methods for quantizing diffusion models, with a taxonomy of post-training and quantization-aware approaches and an analysis of quantization artifacts.
Discussion (0). Continue with ORCID to comment.