Pith. sign in

REVIEW 1 cited by

Deep Contextual Video Compression

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2109.15047 v2 pith:K5Z3PMVX submitted 2021-09-30 eess.IV cs.CVcs.MM

classification eess.IVcs.CVcs.MM
keywords compressionvideocodingdeepframeworkconditionpredictiveconditional
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Most of the existing neural video compression methods adopt the predictive coding framework, which first generates the predicted frame and then encodes its residue with the current frame. However, as for compression ratio, predictive coding is only a sub-optimal solution as it uses simple subtraction operation to remove the redundancy across frames. In this paper, we propose a deep contextual video compression framework to enable a paradigm shift from predictive coding to conditional coding. In particular, we try to answer the following questions: how to define, use, and learn condition under a deep video compression framework. To tap the potential of conditional coding, we propose using feature domain context as condition. This enables us to leverage the high dimension context to carry rich information to both the encoder and the decoder, which helps reconstruct the high-frequency contents for higher video quality. Our framework is also extensible, in which the condition can be flexibly designed. Experiments show that our method can significantly outperform the previous state-of-the-art (SOTA) deep video compression methods. When compared with x265 using veryslow preset, we can achieve 26.0% bitrate saving for 1080P standard test videos.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission

    eess.IV 2024-11 conditional novelty 5.0 of 10

    A cross-layer framework integrates Deep JSCC semantic coding with encryption, CRC, LDPC, and retransmission, adapting error protection to semantic importance for panoramic video.

Pith tools