Pith. sign in

REVIEW 3 cited by

Learning Enriched Features for Real Image Restoration and Enhancement

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2003.06792 v2 pith:4QX5XX7I submitted 2020-03-15 cs.CV

classification cs.CV
keywords imageinformationcontextualfeaturesmulti-scalerepresentationsrestorationachieved
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

With the goal of recovering high-quality image content from its degraded version, image restoration enjoys numerous applications, such as in surveillance, computational photography, medical imaging, and remote sensing. Recently, convolutional neural networks (CNNs) have achieved dramatic improvements over conventional approaches for image restoration task. Existing CNN-based methods typically operate either on full-resolution or on progressively low-resolution representations. In the former case, spatially precise but contextually less robust results are achieved, while in the latter case, semantically reliable but spatially less accurate outputs are generated. In this paper, we present a novel architecture with the collective goals of maintaining spatially-precise high-resolution representations through the entire network and receiving strong contextual information from the low-resolution representations. The core of our approach is a multi-scale residual block containing several key elements: (a) parallel multi-resolution convolution streams for extracting multi-scale features, (b) information exchange across the multi-resolution streams, (c) spatial and channel attention mechanisms for capturing contextual information, and (d) attention based multi-scale feature aggregation. In a nutshell, our approach learns an enriched set of features that combines contextual information from multiple scales, while simultaneously preserving the high-resolution spatial details. Extensive experiments on five real image benchmark datasets demonstrate that our method, named as MIRNet, achieves state-of-the-art results for a variety of image processing tasks, including image denoising, super-resolution, and image enhancement. The source code and pre-trained models are available at https://github.com/swz30/MIRNet.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Wavelet-based Decoupling Framework for low-light Stereo Image Enhancement

    cs.CV 2025-07 conditional novelty 6.0 of 10

    A wavelet-based decoupling network for low-light stereo enhancement, using a low-frequency branch for illumination and high-frequency branches for texture with cross-view interaction, reports state-of-the-art PSNR/SSI...

  2. SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns

    eess.AS 2026-03 conditional novelty 5.5 of 10

    SEMamba++ combines Frequency GLP (FAN-based global-periodic + local conv) with multi-resolution parallel TFDP and learnable softplus mapping to outperform GSR baselines on VCTK, URGENT and AATC while remaining efficient.

  3. Row-Column Separated Attention Based Low-Light Image/Video Enhancement

    cs.CV 2026-02 conditional novelty 5.0 of 10

    A U-Net enhanced with a row/column mean-max attention module reports top PSNR/SSIM on LOL and SDSD while using a fraction of the parameters of transformer rivals.

Pith tools