Pith. sign in

JND-Based Perceptual Optimization For Learned Image Compression

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Recently, learned image compression schemes have achieved remarkable improvements in image fidelity (e.g., PSNR and MS-SSIM) compared to conventional hybrid image coding ones due to their high-efficiency non-linear transform, end-to-end optimization frameworks, etc. However, few of them take the Just Noticeable Difference (JND) characteristic of the Human Visual System (HVS) into account and optimize learned image compression towards perceptual quality. To address this issue, a JND-based perceptual quality loss is proposed. Considering that the amounts of distortion in the compressed image at different training epochs under different Quantization Parameters (QPs) are different, we develop a distortion-aware adjustor. After combining them together, we can better assign the distortion in the compressed image with the guidance of JND to preserve the high perceptual quality. All these designs enable the proposed method to be flexibly applied to various learned image compression schemes with high scalability and plug-and-play advantages. Experimental results on the Kodak dataset demonstrate that the proposed method has led to better perceptual quality than the baseline model under the same bit rate.

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Customizable ROI-Based Deep Image Compression

cs.CV · 2025-07-01 · conditional · novelty 6.0

A text-prompted ROI image coder with a user-controlled quality knob and latent mask attention achieves strong RD and machine-vision results, though headline numbers use ground-truth masks.

citing papers explorer

Showing 1 of 1 citing paper.

  • Customizable ROI-Based Deep Image Compression cs.CV · 2025-07-01 · conditional · none · ref 16 · internal anchor

    A text-prompted ROI image coder with a user-controlled quality knob and latent mask attention achieves strong RD and machine-vision results, though headline numbers use ground-truth masks.