A furniture-removal pipeline for indoor meshes and panoramas that uses a simplified defurnished mesh to guide ControlNet inpainting, producing cleaner results than NeRF or RGB-D baselines.
Self-Supervised Point Cloud Completion via Inpainting
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
When navigating in urban environments, many of the objects that need to be tracked and avoided are heavily occluded. Planning and tracking using these partial scans can be challenging. The aim of this work is to learn to complete these partial point clouds, giving us a full understanding of the object's geometry using only partial observations. Previous methods achieve this with the help of complete, ground-truth annotations of the target objects, which are available only for simulated datasets. However, such ground truth is unavailable for real-world LiDAR data. In this work, we present a self-supervised point cloud completion algorithm, PointPnCNet, which is trained only on partial scans without assuming access to complete, ground-truth annotations. Our method achieves this via inpainting. We remove a portion of the input data and train the network to complete the missing region. As it is difficult to determine which regions were occluded in the initial cloud and which were synthetically removed, our network learns to complete the full cloud, including the missing regions in the initial partial cloud. We show that our method outperforms previous unsupervised and weakly-supervised methods on both the synthetic dataset, ShapeNet, and real-world LiDAR dataset, Semantic KITTI.
citation-role summary
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Defurnishing with X-Ray Vision: Joint Removal of Furniture from Panoramas and Mesh
A furniture-removal pipeline for indoor meshes and panoramas that uses a simplified defurnished mesh to guide ControlNet inpainting, producing cleaner results than NeRF or RGB-D baselines.