Pith. sign in

REVIEW 1 cited by

Tree-NeRV: A Tree-Structured Neural Representation for Efficient Non-Uniform Video Encoding

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.12899 v1 pith:SIJWASJP submitted 2025-04-17 cs.CV

classification cs.CV
keywords samplingtemporaltree-nervvideorepresentationalongaxisefficient
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Implicit Neural Representations for Videos (NeRV) have emerged as a powerful paradigm for video representation, enabling direct mappings from frame indices to video frames. However, existing NeRV-based methods do not fully exploit temporal redundancy, as they rely on uniform sampling along the temporal axis, leading to suboptimal rate-distortion (RD) performance. To address this limitation, we propose Tree-NeRV, a novel tree-structured feature representation for efficient and adaptive video encoding. Unlike conventional approaches, Tree-NeRV organizes feature representations within a Binary Search Tree (BST), enabling non-uniform sampling along the temporal axis. Additionally, we introduce an optimization-driven sampling strategy, dynamically allocating higher sampling density to regions with greater temporal variation. Extensive experiments demonstrate that Tree-NeRV achieves superior compression efficiency and reconstruction quality, outperforming prior uniform sampling-based methods. Code will be released.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SIEDD: Shared-Implicit Encoder with Discrete Decoders

    cs.CV 2025-06 conditional novelty 6.0 of 10

    A shared encoder trained on a few video frames, followed by frozen-encoder parallel decoder training, cuts neural video encoding time by 20 to 30 times at similar quality.

Pith tools