Pith. sign in

REVIEW 7 cited by

Denoising Diffusion Probabilistic Models for Robust Image Super-Resolution in the Wild

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.07864 v1 pith:I67BQRBR submitted 2023-02-15 cs.CV eess.IV

classification cs.CVeess.IV
keywords modelssuper-resolutiontrainingblinddegradationsdiffusionfurtherlarge-scale
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Diffusion models have shown promising results on single-image super-resolution and other image- to-image translation tasks. Despite this success, they have not outperformed state-of-the-art GAN models on the more challenging blind super-resolution task, where the input images are out of distribution, with unknown degradations. This paper introduces SR3+, a diffusion-based model for blind super-resolution, establishing a new state-of-the-art. To this end, we advocate self-supervised training with a combination of composite, parameterized degradations for self-supervised training, and noise-conditioing augmentation during training and testing. With these innovations, a large-scale convolutional architecture, and large-scale datasets, SR3+ greatly outperforms SR3. It outperforms Real-ESRGAN when trained on the same data, with a DRealSR FID score of 36.82 vs. 37.22, which further improves to FID of 32.37 with larger models, and further still with larger training sets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Text-Aware Image Restoration with Diffusion Models

    cs.CV 2025-06 conditional novelty 6.0 of 10

    A diffusion restoration model jointly trained with a text-spotting module and prompted by its own recognized text improves text recognition accuracy on restored images compared with general-purpose restoration methods.

  2. LatentINDIGO: An INN-Guided Latent Diffusion Algorithm for Image Restoration

    cs.CV 2025-05 conditional novelty 6.0 of 10

    LatentINDIGO guides latent diffusion sampling with wavelet-inspired invertible networks, achieving state-of-the-art blind image restoration without retraining the diffusion model.

  3. Robust Knowledge Graph Embedding via Denoising

    cs.LG 2025-05 reject novelty 6.0 of 10

    A denoising auxiliary loss on perturbed entity embeddings improves KGE robustness on FB15k-237, but the proposed certified robustness metrics rest on a misapplication of randomized smoothing.

  4. INDIGO+: A Unified INN-Guided Probabilistic Diffusion Algorithm for Blind and Non-Blind Image Restoration

    cs.CV 2025-01 conditional novelty 6.0 of 10

    An invertible neural network trained to mimic image degradations is used to steer a pretrained diffusion model in every sampling step, giving a blind and non-blind image restoration algorithm.

  5. Adaptive Dropout: Unleashing Dropout across Layers for Generalizable Image Super-Resolution

    cs.CV 2025-06 conditional novelty 5.0 of 10

    A weighted, layer-wise annealed dropout applied at intermediate layers of blind super-resolution networks improves generalization on unseen degradations over prior regularization methods.

  6. NTIRE 2025 Challenge on Short-form UGC Video Quality Assessment and Enhancement: KwaiSR Dataset and Study

    cs.CV 2025-04 conditional novelty 5.0 of 10

    KwaiSR is a new image super-resolution benchmark made from short-form user-generated content, and existing SR models and quality metrics struggle on it.

  7. SupResDiffGAN a new approach for the Super-Resolution task

    eess.IV 2025-04 conditional novelty 4.0 of 10

    SupResDiffGAN combines latent-space diffusion with adversarial training and adaptive input noise, achieving faster super-resolution inference than SR3 and I2SB at comparable LPIPS quality.

Pith tools