Pith. sign in

CANeRV: Content Adaptive Neural Representation for Video Compression

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Recent advances in video compression introduce implicit neural representation (INR) based methods, which effectively capture global dependencies and characteristics of entire video sequences. Unlike traditional and deep learning based approaches, INR-based methods optimize network parameters from a global perspective, resulting in superior compression potential. However, most current INR methods utilize a fixed and uniform network architecture across all frames, limiting their adaptability to dynamic variations within and between video sequences. This often leads to suboptimal compression outcomes as these methods struggle to capture the distinct nuances and transitions in video content. To overcome these challenges, we propose Content Adaptive Neural Representation for Video Compression (CANeRV), an innovative INR-based video compression network that adaptively conducts structure optimisation based on the specific content of each video sequence. To better capture dynamic information across video sequences, we propose a dynamic sequence-level adjustment (DSA). Furthermore, to enhance the capture of dynamics between frames within a sequence, we implement a dynamic frame-level adjustment (DFA). {Finally, to effectively capture spatial structural information within video frames, thereby enhancing the detail restoration capabilities of CANeRV, we devise a structure level hierarchical structural adaptation (HSA).} Experimental results demonstrate that CANeRV can outperform both H.266/VVC and state-of-the-art INR-based video compression techniques across diverse video datasets.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion

cs.CV · 2025-06-18 · conditional · novelty 6.0

MSNeRV is an implicit neural representation video codec that combines temporal-window fusion, GoP-level background grids, multi-resolution supervision, and multi-scale feature blocks, reporting strong compression results on HEVC ClassB and UVG.

citing papers explorer

Showing 1 of 1 citing paper.

  • MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion cs.CV · 2025-06-18 · conditional · none · ref 14 · internal anchor

    MSNeRV is an implicit neural representation video codec that combines temporal-window fusion, GoP-level background grids, multi-resolution supervision, and multi-scale feature blocks, reporting strong compression results on HEVC ClassB and UVG.