Pith. sign in

REVIEW 4 cited by

Block-NeRF: Scalable Large Scene Neural View Synthesis

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2202.05263 v1 pith:V2WWYVDJ submitted 2022-02-10 cs.CV cs.GR

classification cs.CVcs.GR
keywords scenenerfneuralrenderingappearanceblock-nerfenvironmentslarge
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present Block-NeRF, a variant of Neural Radiance Fields that can represent large-scale environments. Specifically, we demonstrate that when scaling NeRF to render city-scale scenes spanning multiple blocks, it is vital to decompose the scene into individually trained NeRFs. This decomposition decouples rendering time from scene size, enables rendering to scale to arbitrarily large environments, and allows per-block updates of the environment. We adopt several architectural changes to make NeRF robust to data captured over months under different environmental conditions. We add appearance embeddings, learned pose refinement, and controllable exposure to each individual NeRF, and introduce a procedure for aligning appearance between adjacent NeRFs so that they can be seamlessly combined. We build a grid of Block-NeRFs from 2.8 million images to create the largest neural scene representation to date, capable of rendering an entire neighborhood of San Francisco.

Discussion (0). Sign in to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language

    cs.CV 2022-04 unverdicted novelty 7.0 of 10

    Socratic Models compose zero-shot multimodal reasoning by prompting pretrained language and vision models to exchange information and enable new capabilities without finetuning.

  2. DiskChunGS: Large-Scale 3D Gaussian SLAM Through Chunk-Based Memory Management

    cs.RO 2025-11 conditional novelty 5.0 of 10

    Storing inactive spatial chunks of a 3D Gaussian map on disk and loading only camera-visible chunks into GPU memory lets DiskChunGS map all 11 KITTI sequences on a 24 GB GPU without memory failures.

  3. Hierarchical vs. Flat Iteration in Shared-Weight Transformers

    cs.CL 2026-04 unverdicted novelty 4.0 of 10

    Hierarchical two-speed shared-weight recurrence in Transformers shows a sharp performance gap compared to independent layer stacking in empirical language modeling tests.

  4. Real-Time Scene Reconstruction using Light Field Probes

    cs.GR 2025-07 conditional novelty 4.0 of 10

    A probe-based renderer built from laser point clouds reconstructs a room-scale scene in real time with constant per-frame cost.

Pith tools