REVIEW 10 cited by
DiffBIR: Towards Blind Image Restoration with Generative Diffusion Prior
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We present DiffBIR, a general restoration pipeline that could handle different blind image restoration tasks in a unified framework. DiffBIR decouples blind image restoration problem into two stages: 1) degradation removal: removing image-independent content; 2) information regeneration: generating the lost image content. Each stage is developed independently but they work seamlessly in a cascaded manner. In the first stage, we use restoration modules to remove degradations and obtain high-fidelity restored results. For the second stage, we propose IRControlNet that leverages the generative ability of latent diffusion models to generate realistic details. Specifically, IRControlNet is trained based on specially produced condition images without distracting noisy content for stable generation performance. Moreover, we design a region-adaptive restoration guidance that can modify the denoising process during inference without model re-training, allowing users to balance realness and fidelity through a tunable guidance scale. Extensive experiments have demonstrated DiffBIR's superiority over state-of-the-art approaches for blind image super-resolution, blind face restoration and blind image denoising tasks on both synthetic and real-world datasets. The code is available at https://github.com/XPixelGroup/DiffBIR.
Forward citations
Cited by 10 Pith papers
-
The Devil is in the Dark Pixels: Toward Brightness Bias-Robust Denoising
BBRD, a drop-in replacement for MSE that normalizes per-brightness-band errors by their own noise variance and upweights the lagging band via softmax Group-DRO, improves dark, bright, and aggregate PSNR simultaneously...
-
MetaScope: Optics-Driven Neural Network for Ultra-Micro Metalens Endoscopy
MetaScope, an optics-driven network, corrects metalens endoscope images and outperforms prior methods on segmentation and restoration.
-
Fine-structure Preserved Real-world Image Super-resolution via Transfer VAE Training
A transfer training scheme converts Stable Diffusion's 8x VAE into a 4x VAE that stays compatible with the pretrained UNet, improving fine-structure preservation in real-world super-resolution at lower FLOPs.
-
Exploiting Gaussian Agnostic Representation Learning with Diffusion Priors for Enhanced Infrared Small Target Detection
Gaussian-sampled non-uniform quantization plus a two-stage diffusion reconstruction generates synthetic infrared images that improve small target detection under data scarcity.
-
Robust ID-Specific Face Restoration via Alignment Learning
RIDFR injects a reference person's identity into diffusion-based face restoration and uses Alignment Learning across multiple same-identity references to suppress pose, expression, and makeup interference.
-
MicroZoom: Structure-Preserving Detail Synthesis at Extreme Scale
A cascaded, segmentation-conditioned, per-instance diffusion method synthesizes globally coherent gigapixel microscopic detail from a phone photo and sparse microscope references at up to 350×.
-
Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration
Pref-Restore combines AR semantic tokens, a diffusion generator, and DiffusionNFT-style RL to make blind face restoration more consistent, but its deterministic-identity claim is weakened by self-referential rewards a...
-
Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution
DDR-SR routes each real-world low-resolution image to one of two diffusion experts based on a high-frequency-loss difficulty score, using a low-compression VAE for hard images and a high-compression VAE for easy image...
-
Teacher-Guided Causal Interventions for Image Denoising: Orthogonal Content-Noise Disentanglement in Vision Transformers
TCD-Net couples a ViT denoiser with "causal" regularizers (de-centering, orthogonality, Nano Banana Pro distillation) and reports marginal PSNR shifts of ≤0.08 dB on some benchmarks, with no error bars or code.
-
Incorporating Uncertainty-Guided and Top-k Codebook Matching for Real-World Blind Image Super-Resolution
UGTSR improves codebook-based blind super-resolution by combining uncertainty-guided loss weighting, top-3 codebook matching, and an align-attention module for fusing low- and high-quality features.
Discussion (0). Sign in to comment.