REVIEW 2 cited by
PUPS: Point Cloud Unified Panoptic Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Point cloud panoptic segmentation is a challenging task that seeks a holistic solution for both semantic and instance segmentation to predict groupings of coherent points. Previous approaches treat semantic and instance segmentation as surrogate tasks, and they either use clustering methods or bounding boxes to gather instance groupings with costly computation and hand-crafted designs in the instance segmentation task. In this paper, we propose a simple but effective point cloud unified panoptic segmentation (PUPS) framework, which use a set of point-level classifiers to directly predict semantic and instance groupings in an end-to-end manner. To realize PUPS, we introduce bipartite matching to our training pipeline so that our classifiers are able to exclusively predict groupings of instances, getting rid of hand-crafted designs, e.g. anchors and Non-Maximum Suppression (NMS). In order to achieve better grouping results, we utilize a transformer decoder to iteratively refine the point classifiers and develop a context-aware CutMix augmentation to overcome the class imbalance problem. As a result, PUPS achieves 1st place on the leader board of SemanticKITTI panoptic segmentation task and state-of-the-art results on nuScenes.
Forward citations
Cited by 2 Pith papers
-
Radar Tracker: Moving Instance Tracking in Sparse and Noisy Radar Point Clouds
Radar Tracker adds temporal offset prediction and attention-based appearance association to a radar instance segmentation backbone, achieving an LSTQ of 66.8 on the RadarScenes moving-instance tracking benchmark.
-
How Do Images Align and Complement LiDAR? Towards a Harmonized Multi-modal 3D Panoptic Segmentation
IAL combines synchronized LiDAR-image augmentation, geometry-guided token fusion, and modality-prior queries to reach state-of-the-art 3D panoptic segmentation on nuScenes and SemanticKITTI.
Discussion (0). Continue with ORCID to comment.