REVIEW 3 cited by
Cylindrical and Asymmetrical 3D Convolution Networks for LiDAR Segmentation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
State-of-the-art methods for large-scale driving-scene LiDAR segmentation often project the point clouds to 2D space and then process them via 2D convolution. Although this corporation shows the competitiveness in the point cloud, it inevitably alters and abandons the 3D topology and geometric relations. A natural remedy is to utilize the3D voxelization and 3D convolution network. However, we found that in the outdoor point cloud, the improvement obtained in this way is quite limited. An important reason is the property of the outdoor point cloud, namely sparsity and varying density. Motivated by this investigation, we propose a new framework for the outdoor LiDAR segmentation, where cylindrical partition and asymmetrical 3D convolution networks are designed to explore the 3D geometric pat-tern while maintaining these inherent properties. Moreover, a point-wise refinement module is introduced to alleviate the interference of lossy voxel-based label encoding. We evaluate the proposed model on two large-scale datasets, i.e., SemanticKITTI and nuScenes. Our method achieves the 1st place in the leaderboard of SemanticKITTI and outperforms existing methods on nuScenes with a noticeable margin, about 4%. Furthermore, the proposed 3D framework also generalizes well to LiDAR panoptic segmentation and LiDAR 3D detection.
Forward citations
Cited by 3 Pith papers
-
FastPoint: Accelerating 3D Point Cloud Model Inference via Sample Point Distance Prediction
FastPoint accelerates farthest point sampling and neighbor search by predicting the FPS minimum distance curve from its first 10%, achieving 2.55x end-to-end speedup with negligible accuracy loss.
-
How Do Images Align and Complement LiDAR? Towards a Harmonized Multi-modal 3D Panoptic Segmentation
IAL combines synchronized LiDAR-image augmentation, geometry-guided token fusion, and modality-prior queries to reach state-of-the-art 3D panoptic segmentation on nuScenes and SemanticKITTI.
-
LiDAR Based Semantic Perception for Forklifts in Outdoor Environments
A dual-LiDAR spherical-projection CNN segments forklift scenes at 74% mIoU with 31 ms inference on an RTX 3090, but the claimed benefit of the second sensor is not ablated.
Discussion (0). Continue with ORCID to comment.