Pith. sign in

REVIEW 22 cited by

NeRF--: Neural Radiance Fields Without Known Camera Parameters

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2102.07064 v4 pith:IZQGA3CA submitted 2021-02-14 cs.CV

classification cs.CV
keywords cameraparametersnerfnoveltrainingviewdatasetforward-facing
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Considering the problem of novel view synthesis (NVS) from only a set of 2D images, we simplify the training process of Neural Radiance Field (NeRF) on forward-facing scenes by removing the requirement of known or pre-computed camera parameters, including both intrinsics and 6DoF poses. To this end, we propose NeRF$--$, with three contributions: First, we show that the camera parameters can be jointly optimised as learnable parameters with NeRF training, through a photometric reconstruction; Second, to benchmark the camera parameter estimation and the quality of novel view renderings, we introduce a new dataset of path-traced synthetic scenes, termed as Blender Forward-Facing Dataset (BLEFF); Third, we conduct extensive analyses to understand the training behaviours under various camera motions, and show that in most scenarios, the joint optimisation pipeline can recover accurate camera parameters and achieve comparable novel view synthesis quality as those trained with COLMAP pre-computed camera parameters. Our code and data are available at https://nerfmm.active.vision.

Discussion (0). Sign in to comment.

Forward citations

Cited by 22 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. CalibAnyView: Beyond Single-View Camera Calibration in the Wild

    cs.CV 2026-05 conditional novelty 8.0 of 10

    A multi-view transformer predicts dense perspective fields that feed a geometric optimizer to estimate camera intrinsics and gravity from arbitrary numbers of real-world views.

  2. HairGPT: Strand-as-Language Autoregressive Modeling for Realistic 3D Hairstyle Synthesis

    cs.GR 2026-05 unverdicted novelty 7.0 of 10

    HairGPT reframes 3D hairstyle synthesis as dual-decoupled autoregressive strand sequence modeling with geometric tokenization for semantic control and rare style generation.

  3. NoDrift3R: Raymap-Guided Coupling for Drift-Robust Unposed Feed-Forward 3D Reconstruction

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Anchoring Gaussian centers to predicted raymaps and jointly optimizing RGB, raymap, and camera losses with a dual-frequency curriculum suppresses pose drift and improves pose-free 3D reconstruction on long sequences.

  4. StructSplat: Generalizable 3D Gaussian Splatting from Uncalibrated Sparse Views

    cs.CV 2026-06 unverdicted novelty 6.0 of 10

    StructSplat introduces a structured 3D Gaussian splatting framework that performs feed-forward reconstruction from uncalibrated sparse views using pixel-aligned features, semantic priors, and camera alignment.

  5. RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging

    cs.CV 2026-04 unverdicted novelty 6.0 of 10

    RayFormer improves NeRF reconstruction for video SCI by replacing random ray sampling with patch-level sampling, adding a transformer to capture inter- and intra-ray structural similarities, and incorporating a total ...

  6. PCM-NeRF: Probabilistic Camera Modeling for Neural Radiance Fields under Pose Uncertainty

    cs.CV 2026-04 unverdicted novelty 6.0 of 10

    PCM-NeRF improves neural surface reconstruction under uncertain camera poses by learning per-camera pose distributions and damping updates from high-uncertainty views.

  7. LiveStre4m: Feed-Forward Live Streaming of Novel Views from Unposed Multi-View Video

    cs.CV 2026-04 unverdicted novelty 6.0 of 10

    LiveStre4m delivers real-time novel-view video streaming from unposed multi-view inputs via a multi-view vision transformer, diffusion-transformer interpolation, and a learned camera pose predictor.

  8. LongSplat: Robust Unposed 3D Gaussian Splatting for Casual Long Videos

    cs.CV 2025-08 conditional novelty 6.0 of 10

    An incremental 3D Gaussian Splatting pipeline that jointly optimizes camera poses and scene geometry using MASt3R priors and density-adaptive octree anchors achieves state-of-the-art novel view synthesis on casual lon...

  9. The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with Minimal 3D Knowledge

    cs.CV 2025-06 unverdicted novelty 6.0 of 10

    Data-centric novel view synthesis models with minimal 3D knowledge and no pose annotations scale better with data volume and outperform traditional bias-driven methods.

  10. RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos

    cs.CV 2024-12 unverdicted novelty 6.0 of 10

    RoDyGS separates static and dynamic elements in monocular videos using Gaussian splatting with regularization and introduces the Kubric-MRig benchmark for pose-free dynamic novel view synthesis.

  11. SalientGS: Unified SfM-to-3DGS with Importance-Guided MCMC Gaussian Allocation

    cs.CV 2026-07 conditional novelty 5.5 of 10

    Importance-guided MCMC reallocates 3D Gaussians toward multi-view underfit regions, enabling a unified SfM-to-3DGS pipeline that finishes in ~15 minutes with SOTA perceptual quality.

  12. SalientGS: Unified SfM-to-3DGS with Importance-Guided MCMC Gaussian Allocation

    cs.CV 2026-07 conditional novelty 5.0 of 10

    SalientGS integrates fast first-order SfM, joint pose refinement, and importance-guided MCMC Gaussian birth/relocation to reach 27.65 dB macro-average PSNR at 1.5M Gaussians in about 10 minutes end-to-end.

  13. NoDrift3R: Raymap-Guided Coupling for Drift-Robust Unposed Feed-Forward 3D Reconstruction

    cs.CV 2026-07 conditional novelty 5.0 of 10

    Anchoring 3D Gaussian centers to ray-map predictions and jointly optimizing geometry with appearance supervision suppresses pose drift in unposed feed-forward 3D reconstruction.

  14. MZEN: Multi-Zoom Enhanced NeRF for 3-D Reconstruction with Unknown Camera Poses

    cs.CV 2025-08 reject novelty 5.0 of 10

    MZEN is a NeRF training schedule that handles multi-zoom image sets with a zoom-scaled camera model and a bootstrap-register-refine pose strategy, claiming large gains on a new eight-scene benchmark.

  15. KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos

    cs.CV 2024-11 unverdicted novelty 5.0 of 10

    KFC-W is a self-supervised 3D-aware video model trained on videos and multiview internet photos that produces geometrically consistent interpolations between unposed input images without any 3D annotations.

  16. Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs

    cs.CV 2024-08 unverdicted novelty 5.0 of 10

    Splatt3R is a feed-forward network that predicts 3D Gaussian splats directly from uncalibrated stereo image pairs by extending MASt3R with appearance attributes and a two-stage training procedure.

  17. BSNeRF: Broadband Spectral Neural Radiance Fields for Snapshot Multispectral Light-field Imaging

    eess.SP 2025-09 reject novelty 4.0 of 10

    A neural radiance field with joint camera pose estimation is proposed to decouple broadband spectral multiplexing in snapshot multispectral light-field images.

  18. Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning

    cs.RO 2025-08 conditional novelty 4.0 of 10

    The thesis demonstrates that combining implicit 3D scene representations with LLM-based reasoning, using text as an interface, yields strong performance on robotic perception and spatial language tasks.

  19. DrivingGaussian++: Towards Realistic Reconstruction and Editable Simulation for Surrounding Dynamic Driving Scenes

    cs.CV 2025-08 conditional novelty 4.0 of 10

    DrivingGaussian++ reconstructs dynamic surround-view driving scenes and performs training-free multi-task editing (weather, texture, object manipulation) using Gaussians, diffusion models, and LLM-generated trajectories.

  20. Novel View Synthesis with Gaussian Splatting: Impact on Photogrammetry Model Accuracy and Resolution

    cs.CV 2025-08 conditional novelty 4.0 of 10

    Gaussian Splatting beats photogrammetry on image quality metrics, and adding its synthetic views to a photogrammetry dataset raises SSIM/PSNR but lowers measured resolution.

  21. Neural Field Representations of Mobile Computational Photography

    cs.CV 2025-08 conditional novelty 4.0 of 10

    Fitting neural fields directly to raw phone bursts reconstructs depth, separates reflections and occluders, and stitches panoramas, outperforming the compared baselines on the thesis's benchmarks.

  22. NeRF: Neural Radiance Field in 3D Vision: A Comprehensive Review (Updated Post-Gaussian Splatting)

    cs.CV 2022-10 unverdicted novelty 2.0 of 10

    A literature survey of NeRF and neural field methods from 2020-2025, organized by architecture and application taxonomies with benchmarks and dataset overviews, covering both pre- and post-Gaussian Splatting periods.

Pith tools