Pith. sign in

REVIEW 2 cited by

Transfer CLIP for Generalizable Image Denoising

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.15132 v1 pith:UPMGE2EK submitted 2024-03-22 cs.CV eess.IV

classification cs.CVeess.IV
keywords imageclipdenoisingnoisefeaturesgeneralizabledecoderdense
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Image denoising is a fundamental task in computer vision. While prevailing deep learning-based supervised and self-supervised methods have excelled in eliminating in-distribution noise, their susceptibility to out-of-distribution (OOD) noise remains a significant challenge. The recent emergence of contrastive language-image pre-training (CLIP) model has showcased exceptional capabilities in open-world image recognition and segmentation. Yet, the potential for leveraging CLIP to enhance the robustness of low-level tasks remains largely unexplored. This paper uncovers that certain dense features extracted from the frozen ResNet image encoder of CLIP exhibit distortion-invariant and content-related properties, which are highly desirable for generalizable denoising. Leveraging these properties, we devise an asymmetrical encoder-decoder denoising network, which incorporates dense features including the noisy image and its multi-scale features from the frozen ResNet encoder of CLIP into a learnable image decoder to achieve generalizable denoising. The progressive feature augmentation strategy is further proposed to mitigate feature overfitting and improve the robustness of the learnable decoder. Extensive experiments and comparisons conducted across diverse OOD noises, including synthetic noise, real-world sRGB noise, and low-dose CT image noise, demonstrate the superior generalization ability of our method.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Robust Image Denoising with Scale Equivariance

    cs.CV 2025-08 conditional novelty 6.0 of 10

    Scale-equivariant (first-order homogeneous) network components are proposed as an inductive bias that lets denoisers trained on uniform Gaussian noise generalize to spatially varying OOD noise.

  2. Single Domain Generalization for Few-Shot Counting via Universal Representation Matching

    cs.CV 2025-05 conditional novelty 6.0 of 10

    URM distills CLIP vision-language representations into learnable prototypes for few-shot counting, improving single-domain generalization on unseen datasets.

Pith tools