Pith. sign in

Learned Lossless Compression for JPEG via Frequency-Domain Prediction

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

JPEG images can be further compressed to enhance the storage and transmission of large-scale image datasets. Existing learned lossless compressors for RGB images cannot be well transferred to JPEG images due to the distinguishing distribution of DCT coefficients and raw pixels. In this paper, we propose a novel framework for learned lossless compression of JPEG images that achieves end-to-end optimized prediction of the distribution of decoded DCT coefficients. To enable learning in the frequency domain, DCT coefficients are partitioned into groups to utilize implicit local redundancy. An autoencoder-like architecture is designed based on the weight-shared blocks to realize entropy modeling of grouped DCT coefficients and independently compress the priors. We attempt to realize learned lossless compression of JPEG images in the frequency domain. Experimental results demonstrate that the proposed framework achieves superior or comparable performance in comparison to most recent lossless compressors with handcrafted context modeling for JPEG images.

citation-role summary

background 1

citation-polarity summary

fields

eess.IV 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

SMIC: Semantic Multi-Item Compression based on CLIP dictionary

eess.IV · 2024-12-06 · conditional · novelty 6.0

A dictionary-based multi-item codec sparsely projects CLIP image embeddings onto learned semantic atoms and regenerates images with unCLIP, reaching about 1e-4 BPP per image on a 5000-image collection.

citing papers explorer

Showing 1 of 1 citing paper.

  • SMIC: Semantic Multi-Item Compression based on CLIP dictionary eess.IV · 2024-12-06 · conditional · none · ref 9 · internal anchor

    A dictionary-based multi-item codec sparsely projects CLIP image embeddings onto learned semantic atoms and regenerates images with unCLIP, reaching about 1e-4 BPP per image on a 5000-image collection.