REVIEW 2 cited by
DPFT: Dual Perspective Fusion Transformer for Camera-Radar-based Object Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The perception of autonomous vehicles has to be efficient, robust, and cost-effective. However, cameras are not robust against severe weather conditions, lidar sensors are expensive, and the performance of radar-based perception is still inferior to the others. Camera-radar fusion methods have been proposed to address this issue, but these are constrained by the typical sparsity of radar point clouds and often designed for radars without elevation information. We propose a novel camera-radar fusion approach called Dual Perspective Fusion Transformer (DPFT), designed to overcome these limitations. Our method leverages lower-level radar data (the radar cube) instead of the processed point clouds to preserve as much information as possible and employs projections in both the camera and ground planes to effectively use radars with elevation information and simplify the fusion with camera data. As a result, DPFT has demonstrated state-of-the-art performance on the K-Radar dataset while showing remarkable robustness against adverse weather conditions and maintaining a low inference time. The code is made available as open-source software under https://github.com/TUMFTM/DPFT.
Forward citations
Cited by 2 Pith papers
-
CVFusion: Cross-View Fusion of 4D Radar and Camera for 3D Object Detection
A two-stage cross-view fusion network for 4D radar and camera achieves new state-of-the-art 3D detection accuracy on VoD and TJ4DRadSet, and is the first radar-camera fusion method to match a LiDAR-based detector on VoD.
-
4DR P2T: 4D Radar Tensor Synthesis with Point Clouds
A conditional GAN with 3D sparse and dense convolutions reconstructs 4D radar tensor volumes from point clouds, with a percentile-based comparison showing 1% density balances compression and quality.
Discussion (0). Continue with ORCID to comment.