Pith. sign in

REVIEW 3 cited by

Diffusion Models in Low-Level Vision: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.11138 v2 pith:23K7EFRY submitted 2024-06-17 cs.CV cs.AI

classification cs.CVcs.AI
keywords diffusionmodelsmodel-basedtaskslow-levelvisioncomprehensivegenerative
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep generative models have garnered significant attention in low-level vision tasks due to their generative capabilities. Among them, diffusion model-based solutions, characterized by a forward diffusion process and a reverse denoising process, have emerged as widely acclaimed for their ability to produce samples of superior quality and diversity. This ensures the generation of visually compelling results with intricate texture information. Despite their remarkable success, a noticeable gap exists in a comprehensive survey that amalgamates these pioneering diffusion model-based works and organizes the corresponding threads. This paper proposes the comprehensive review of diffusion model-based techniques. We present three generic diffusion modeling frameworks and explore their correlations with other deep generative models, establishing the theoretical foundation. Following this, we introduce a multi-perspective categorization of diffusion models, considering both the underlying framework and the target task. Additionally, we summarize extended diffusion models applied in other tasks, including medical, remote sensing, and video scenarios. Moreover, we provide an overview of commonly used benchmarks and evaluation metrics. We conduct a thorough evaluation, encompassing both performance and efficiency, of diffusion model-based techniques in three prominent tasks. Finally, we elucidate the limitations of current diffusion models and propose seven intriguing directions for future research. This comprehensive examination aims to facilitate a profound understanding of the landscape surrounding denoising diffusion models in the context of low-level vision tasks. A curated list of diffusion model-based techniques in over 20 low-level vision tasks can be found at https://github.com/ChunmingHe/awesome-diffusion-models-in-low-level-vision.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Uncertainty-Masked Bernoulli Diffusion for Camouflaged Object Detection Refinement

    cs.CV 2025-06 conditional novelty 6.0 of 10

    An uncertainty-masked Bernoulli diffusion refiner improves camouflaged object detection masks from existing models, achieving average gains of 5.5% in MAE and 3.2% in weighted F-measure.

  2. DarkDiff: Advancing Low-Light Raw Enhancement by Retasking Diffusion Models for Camera ISP

    cs.CV 2025-05 conditional novelty 6.0 of 10

    DarkDiff fine-tunes Stable Diffusion with region-based cross-attention, a residual VAE, and a pixel-space loss to turn noisy linear-RGB low-light images into clean sRGB photos, achieving top LPIPS on SID, ELD, and LRD.

  3. A Synthetic-to-Real Dehazing Method based on Domain Unification

    eess.IV 2025-09 conditional novelty 4.0 of 10

    The paper derives a composite atmospheric scattering model for non-ideal clean data and uses a four-loss committee to train a dehazing network, claiming state-of-the-art results on real-world haze benchmarks.

Pith tools