Pith. sign in

REVIEW 1 cited by

Generalizable Neural Radiance Fields for Novel View Synthesis with Transformer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2206.05375 v1 pith:2CGBGR3B submitted 2022-06-10 cs.CV

classification cs.CV
keywords viewviewsrenderingsourceneuralnovelradiancetransnerf
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We propose a Transformer-based NeRF (TransNeRF) to learn a generic neural radiance field conditioned on observed-view images for the novel view synthesis task. By contrast, existing MLP-based NeRFs are not able to directly receive observed views with an arbitrary number and require an auxiliary pooling-based operation to fuse source-view information, resulting in the missing of complicated relationships between source views and the target rendering view. Furthermore, current approaches process each 3D point individually and ignore the local consistency of a radiance field scene representation. These limitations potentially can reduce their performance in challenging real-world applications where large differences between source views and a novel rendering view may exist. To address these challenges, our TransNeRF utilizes the attention mechanism to naturally decode deep associations of an arbitrary number of source views into a coordinate-based scene representation. Local consistency of shape and appearance are considered in the ray-cast space and the surrounding-view space within a unified Transformer network. Experiments demonstrate that our TransNeRF, trained on a wide variety of scenes, can achieve better performance in comparison to state-of-the-art image-based neural rendering methods in both scene-agnostic and per-scene finetuning scenarios especially when there is a considerable gap between source views and a rendering view.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View Synthesis

    cs.CV 2025-05 conditional novelty 6.0 of 10

    GoLF-NRT fuses global context from a 3D sparse-attention transformer with epipolar local geometry and kernel-regression adaptive sampling to improve few-shot novel view synthesis.

Pith tools