Pith. sign in

Behind the Scenes: Density Fields for Single View Reconstruction

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Inferring a meaningful geometric scene representation from a single image is a fundamental problem in computer vision. Approaches based on traditional depth map prediction can only reason about areas that are visible in the image. Currently, neural radiance fields (NeRFs) can capture true 3D including color, but are too complex to be generated from a single image. As an alternative, we propose to predict implicit density fields. A density field maps every location in the frustum of the input image to volumetric density. By directly sampling color from the available views instead of storing color in the density field, our scene representation becomes significantly less complex compared to NeRFs, and a neural network can predict it in a single forward pass. The prediction network is trained through self-supervision from only video data. Our formulation allows volume rendering to perform both depth prediction and novel view synthesis. Through experiments, we show that our method is able to predict meaningful geometry for regions that are occluded in the input image. Additionally, we demonstrate the potential of our approach on three datasets for depth prediction and novel-view synthesis.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Masks make discriminative models great again!

cs.CV · 2025-07-01 · conditional · novelty 6.0

Training a single-image 3D Gaussian splat model on visible regions only, using visibility masks from optimized per-scene splats, improves reconstruction quality in visible areas and stays competitive with full-scene models.

citing papers explorer

Showing 1 of 1 citing paper.

  • Masks make discriminative models great again! cs.CV · 2025-07-01 · conditional · none · ref 16 · internal anchor

    Training a single-image 3D Gaussian splat model on visible regions only, using visibility masks from optimized per-scene splats, improves reconstruction quality in visible areas and stays competitive with full-scene models.