A self-supervised multimodal NeRF framework jointly learns LiDAR and camera novel view synthesis for static and dynamic driving scenes, beating LiDAR-NeRF and LiDAR4D on KITTI-360.
LidaRF: Delving into Lidar for Neural Radiance Field on Street Scenes
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Photorealistic simulation plays a crucial role in applications such as autonomous driving, where advances in neural radiance fields (NeRFs) may allow better scalability through the automatic creation of digital 3D assets. However, reconstruction quality suffers on street scenes due to largely collinear camera motions and sparser samplings at higher speeds. On the other hand, the application often demands rendering from camera views that deviate from the inputs to accurately simulate behaviors like lane changes. In this paper, we propose several insights that allow a better utilization of Lidar data to improve NeRF quality on street scenes. First, our framework learns a geometric scene representation from Lidar, which is fused with the implicit grid-based representation for radiance decoding, thereby supplying stronger geometric information offered by explicit point cloud. Second, we put forth a robust occlusion-aware depth supervision scheme, which allows utilizing densified Lidar points by accumulation. Third, we generate augmented training views from Lidar points for further improvement. Our insights translate to largely improved novel view synthesis under real driving scenes.
citation-role summary
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Self-Supervised Multimodal NeRF for Autonomous Driving
A self-supervised multimodal NeRF framework jointly learns LiDAR and camera novel view synthesis for static and dynamic driving scenes, beating LiDAR-NeRF and LiDAR4D on KITTI-360.