REVIEW 4 cited by
A Unified End-to-End Framework for Efficient Deep Image Compression
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Image compression is a widely used technique to reduce the spatial redundancy in images. Recently, learning based image compression has achieved significant progress by using the powerful representation ability from neural networks. However, the current state-of-the-art learning based image compression methods suffer from the huge computational cost, which limits their capacity for practical applications. In this paper, we propose a unified framework called Efficient Deep Image Compression (EDIC) based on three new technologies, including a channel attention module, a Gaussian mixture model and a decoder-side enhancement module. Specifically, we design an auto-encoder style network for learning based image compression. To improve the coding efficiency, we exploit the channel relationship between latent representations by using the channel attention module. Besides, the Gaussian mixture model is introduced for the entropy model and improves the accuracy for bitrate estimation. Furthermore, we introduce the decoder-side enhancement module to further improve image compression performance. Our EDIC method can also be readily incorporated with the Deep Video Compression (DVC) framework to further improve the video compression performance. Simultaneously, our EDIC method boosts the coding performance significantly while bringing slightly increased computational cost. More importantly, experimental results demonstrate that the proposed approach outperforms the current state-of-the-art image compression methods and is up to more than 150 times faster in terms of decoding speed when compared with Minnen's method. The proposed framework also successfully improves the performance of the recent deep video compression system DVC. Our code will be released at https://github.com/liujiaheng/compression.
Forward citations
Cited by 4 Pith papers
-
Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors
Ultra-low-bitrate image decoding is cast as one-step next-frame prediction from a compact anchor using adapted video diffusion priors, yielding large perceptual bitrate savings versus DiffC.
-
Perception-Oriented Latent Coding for High-Performance Compressed Domain Semantic Inference
Perception-oriented training of learned image codecs yields latent codes that support high-accuracy classification and segmentation with only a small adapter fine-tuned.
-
S2CFormer: Revisiting the RD-Latency Trade-off in Transformer-based Learned Image Compression
Channel aggregation by feed-forward networks, not spatial attention, drives rate-distortion performance in transformer-based learned image compression, and simplified models achieve state-of-the-art results with over ...
-
Conditional Latent Coding with Learnable Synthesized Reference for Deep Image Compression
Conditional Latent Coding compresses images by synthesizing a per-image reference latent from a learned feature dictionary, improving low-bitrate rate-distortion over TCM, VTM, and BPG.
Discussion (0). Continue with ORCID to comment.