Pith. sign in

REVIEW 1 cited by

PoinTr: Diverse Point Cloud Completion with Geometry-Aware Transformers

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2108.08839 v1 pith:L4YMR45E submitted 2021-08-19 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords pointcloudcloudscompletiontransformersbetterpointrapplications
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Point clouds captured in real-world applications are often incomplete due to the limited sensor resolution, single viewpoint, and occlusion. Therefore, recovering the complete point clouds from partial ones becomes an indispensable task in many practical applications. In this paper, we present a new method that reformulates point cloud completion as a set-to-set translation problem and design a new model, called PoinTr that adopts a transformer encoder-decoder architecture for point cloud completion. By representing the point cloud as a set of unordered groups of points with position embeddings, we convert the point cloud to a sequence of point proxies and employ the transformers for point cloud generation. To facilitate transformers to better leverage the inductive bias about 3D geometric structures of point clouds, we further devise a geometry-aware block that models the local geometric relationships explicitly. The migration of transformers enables our model to better learn structural knowledge and preserve detailed information for point cloud completion. Furthermore, we propose two more challenging benchmarks with more diverse incomplete point clouds that can better reflect the real-world scenarios to promote future research. Experimental results show that our method outperforms state-of-the-art methods by a large margin on both the new benchmarks and the existing ones. Code is available at https://github.com/yuxumin/PoinTr

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Multimodal Human-Intent Modeling for Contextual Robot-to-Human Handovers of Arbitrary Objects

    cs.RO 2025-08 conditional novelty 5.0 of 10

    A gaze-plus-language pipeline enables a robot to select tabletop objects from a remote user's monitor and generate human-aware grasps for handover, with real-world tests on YCB objects.

Pith tools