REVIEW 2 cited by
CCD-3DR: Consistent Conditioning in Diffusion for Single-Image 3D Reconstruction
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper, we present a novel shape reconstruction method leveraging diffusion model to generate 3D sparse point cloud for the object captured in a single RGB image. Recent methods typically leverage global embedding or local projection-based features as the condition to guide the diffusion model. However, such strategies fail to consistently align the denoised point cloud with the given image, leading to unstable conditioning and inferior performance. In this paper, we present CCD-3DR, which exploits a novel centered diffusion probabilistic model for consistent local feature conditioning. We constrain the noise and sampled point cloud from the diffusion model into a subspace where the point cloud center remains unchanged during the forward diffusion process and reverse process. The stable point cloud center further serves as an anchor to align each point with its corresponding local projection-based features. Extensive experiments on synthetic benchmark ShapeNet-R2N2 demonstrate that CCD-3DR outperforms all competitors by a large margin, with over 40% improvement. We also provide results on real-world dataset Pix3D to thoroughly demonstrate the potential of CCD-3DR in real-world applications. Codes will be released soon
Forward citations
Cited by 2 Pith papers
-
SeaLion: Semantic Part-Aware Latent Point Diffusion Models for 3D Generation
A latent diffusion model that jointly generates 3D point clouds and their semantic part segmentations, plus a part-aware Chamfer distance metric for evaluating them.
-
Consistency Diffusion Models for Single-Image 3D Reconstruction with Priors
A diffusion model for single-image 3D reconstruction that adds multi-view depth-projection consistency and DINOv2-derived 2D priors to the PC2 training objective.
Discussion (0). Continue with ORCID to comment.