Pith. sign in

REVIEW 1 cited by

Lossy Image Compression with Foundation Diffusion Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.08580 v2 pith:LPFWO7V6 submitted 2024-04-12 eess.IV cs.CV

classification eess.IVcs.CV
keywords diffusionmodelsimagemethodscompressionfoundationgenerativemodel
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Incorporating diffusion models in the image compression domain has the potential to produce realistic and detailed reconstructions, especially at extremely low bitrates. Previous methods focus on using diffusion models as expressive decoders robust to quantization errors in the conditioning signals, yet achieving competitive results in this manner requires costly training of the diffusion model and long inference times due to the iterative generative process. In this work we formulate the removal of quantization error as a denoising task, using diffusion to recover lost information in the transmitted image latent. Our approach allows us to perform less than 10% of the full diffusion generative process and requires no architectural changes to the diffusion model, enabling the use of foundation models as a strong prior without additional fine tuning of the backbone. Our proposed codec outperforms previous methods in quantitative realism metrics, and we verify that our reconstructions are qualitatively preferred by end users, even when other methods use twice the bitrate.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Diffusion-based Perceptual Neural Video Compression with Temporal Diffusion Information Reuse

    cs.CV 2025-01 conditional novelty 6.0 of 10

    DiffVC integrates Stable Diffusion into a conditional neural video codec, with temporal reuse of diffusion predictions for speed and quantization-parameter prompting for variable bitrate, achieving state-of-the-art pe...

Pith tools