Pith. sign in

REVIEW 4 cited by

NeSF: Neural Semantic Fields for Generalizable Semantic Segmentation of 3D Scenes

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2111.13260 v3 pith:ZJYTSOFB submitted 2021-11-25 cs.CV cs.RO

classification cs.CVcs.RO
keywords semanticmethoddensityfieldsnesfscenessegmentationalone
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present NeSF, a method for producing 3D semantic fields from posed RGB images alone. In place of classical 3D representations, our method builds on recent work in implicit neural scene representations wherein 3D structure is captured by point-wise functions. We leverage this methodology to recover 3D density fields upon which we then train a 3D semantic segmentation model supervised by posed 2D semantic maps. Despite being trained on 2D signals alone, our method is able to generate 3D-consistent semantic maps from novel camera poses and can be queried at arbitrary 3D points. Notably, NeSF is compatible with any method producing a density field, and its accuracy improves as the quality of the density field improves. Our empirical analysis demonstrates comparable quality to competitive 2D and 3D semantic segmentation baselines on complex, realistically rendered synthetic scenes. Our method is the first to offer truly dense 3D scene segmentations requiring only 2D supervision for training, and does not require any semantic input for inference on novel scenes. We encourage the readers to visit the project website.

Discussion (0). Sign in to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. 3AM: 3egment Anything with Geometric Consistency in Videos

    cs.CV 2026-01 unverdicted novelty 7.0 of 10

    3AM integrates MUSt3R 3D features into SAM2 via a Feature Merger and FOV-aware sampling to deliver geometry-consistent video object segmentation from RGB alone, with large gains on wide-baseline datasets.

  2. OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation

    cs.CV 2025-12 unverdicted novelty 7.0 of 10

    OpenTrack3D achieves state-of-the-art open-vocabulary 3D instance segmentation by generating cross-view consistent proposals online with a visual-spatial tracker and replacing CLIP with an MLLM for improved compositio...

  3. First Shape, Then Meaning: Efficient Geometry and Semantics Learning for Indoor Reconstruction

    cs.CV 2026-05 unverdicted novelty 5.0 of 10

    FSTM improves indoor reconstruction by training geometry first without semantic supervision, then adding semantics, achieving 2.3x faster training and higher object surface recall than joint optimization.

  4. A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding

    cs.CV 2025-07 conditional novelty 5.0 of 10

    SpatialReasoner adds LLM-based query decomposition and a visual-properties-enhanced hierarchical feature field to 3D language fields, improving instance localization for spatial-relation queries on self-constructed be...

Pith tools