REVIEW 2 cited by
V2X-PC: Vehicle-to-everything Collaborative Perception via Point Cluster
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The objective of the collaborative vehicle-to-everything perception task is to enhance the individual vehicle's perception capability through message communication among neighboring traffic agents. Previous methods focus on achieving optimal performance within bandwidth limitations and typically adopt BEV maps as the basic collaborative message units. However, we demonstrate that collaboration with dense representations is plagued by object feature destruction during message packing, inefficient message aggregation for long-range collaboration, and implicit structure representation communication. To tackle these issues, we introduce a brand new message unit, namely point cluster, designed to represent the scene sparsely with a combination of low-level structure information and high-level semantic information. The point cluster inherently preserves object information while packing messages, with weak relevance to the collaboration range, and supports explicit structure modeling. Building upon this representation, we propose a novel framework V2X-PC for collaborative perception. This framework includes a Point Cluster Packing (PCP) module to keep object feature and manage bandwidth through the manipulation of cluster point numbers. As for effective message aggregation, we propose a Point Cluster Aggregation (PCA) module to match and merge point clusters associated with the same object. To further handle time latency and pose errors encountered in real-world scenarios, we propose parameter-free solutions that can adapt to different noisy levels without finetuning. Experiments on two widely recognized collaborative perception benchmarks showcase the superior performance of our method compared to the previous state-of-the-art approaches relying on BEV maps.
Forward citations
Cited by 2 Pith papers
-
Is Intermediate Fusion All You Need for UAV-based Collaborative Perception?
A late-intermediate fusion method that transmits only 2D and 3D detection boxes and confidence scores among UAVs, then injects them into the receiver's BEV features, achieves 72.1% mAP on UAV3D with minimal bandwidth.
-
Collaborative Perception Datasets for Autonomous Driving: A Review
A structured survey that catalogs and compares collaborative perception datasets for autonomous driving across cooperation paradigms, sensors, scenarios, and tasks, with a living online repository.
Discussion (0). Continue with ORCID to comment.