Pith. sign in

REVIEW 2 cited by

3DGStream: On-the-Fly Training of 3D Gaussians for Efficient Streaming of Photo-Realistic Free-Viewpoint Videos

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.01444 v4 pith:QFGIOKJS submitted 2024-03-03 cs.CV

classification cs.CV
keywords renderingtrainingdgstreamdynamicscenesvideosachievesefficient
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Constructing photo-realistic Free-Viewpoint Videos (FVVs) of dynamic scenes from multi-view videos remains a challenging endeavor. Despite the remarkable advancements achieved by current neural rendering techniques, these methods generally require complete video sequences for offline training and are not capable of real-time rendering. To address these constraints, we introduce 3DGStream, a method designed for efficient FVV streaming of real-world dynamic scenes. Our method achieves fast on-the-fly per-frame reconstruction within 12 seconds and real-time rendering at 200 FPS. Specifically, we utilize 3D Gaussians (3DGs) to represent the scene. Instead of the na\"ive approach of directly optimizing 3DGs per-frame, we employ a compact Neural Transformation Cache (NTC) to model the translations and rotations of 3DGs, markedly reducing the training time and storage required for each FVV frame. Furthermore, we propose an adaptive 3DG addition strategy to handle emerging objects in dynamic scenes. Experiments demonstrate that 3DGStream achieves competitive performance in terms of rendering speed, image quality, training time, and model storage when compared with state-of-the-art methods.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. QUEEN: QUantized Efficient ENcoding of Dynamic Gaussians for Streaming Free-viewpoint Videos

    cs.CV 2024-12 conditional novelty 6.0 of 10

    QUEEN compresses per-frame Gaussian residuals with learned quantization and gating, reaching about 0.7 MB per frame, under 5 seconds of training, and 350 FPS rendering on dynamic scenes.

  2. Monocular Dynamic Gaussian Splatting: Fast, Brittle, and Scene Complexity Rules

    cs.CV 2024-12 conditional novelty 6.0 of 10

    A comprehensive benchmark shows monocular dynamic Gaussian splatting methods are fast and brittle, with scene complexity dominating method differences.

Pith tools