REVIEW 3 cited by
Wavelet-Like Transform-Based Technology in Response to the Call for Proposals on Neural Network-Based Image Coding
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Neural network-based image coding has been developing rapidly since its birth. Until 2022, its performance has surpassed that of the best-performing traditional image coding framework -- H.266/VVC. Witnessing such success, the IEEE 1857.11 working subgroup initializes a neural network-based image coding standard project and issues a corresponding call for proposals (CfP). In response to the CfP, this paper introduces a novel wavelet-like transform-based end-to-end image coding framework -- iWaveV3. iWaveV3 incorporates many new features such as affine wavelet-like transform, perceptual-friendly quality metric, and more advanced training and online optimization strategies into our previous wavelet-like transform-based framework iWave++. While preserving the features of supporting lossy and lossless compression simultaneously, iWaveV3 also achieves state-of-the-art compression efficiency for objective quality and is very competitive for perceptual quality. As a result, iWaveV3 is adopted as a candidate scheme for developing the IEEE Standard for neural-network-based image coding.
Forward citations
Cited by 3 Pith papers
-
Learning Switchable Priors for Neural Image Compression
A finite set of trainable priors, selected by predicted indices, decouples entropy coding complexity from the probabilistic model family in neural image compression, enabling faster and lighter codecs that still beat BPG.
-
The Gap Between Principle and Practice of Lossy Image Coding
The paper attributes the gap between ideal and practical lossy image coding to five effects, and reports an estimated rate-distortion upper bound that beats VTM by up to 35 percent on Kodak.
-
Generalized Gaussian Model for Learned Image Compression
A generalized Gaussian entropy model with a learned shape parameter and two training fixes improves rate-distortion performance of learned image codecs compared to Gaussian and mixture models.
Discussion (0). Continue with ORCID to comment.