PVCap combines instance-mixing data augmentation with pseudo-labels and a voxel-based captioning network to achieve new state-of-the-art on 3D dense captioning benchmarks ScanRefer and Nr3D.
Fcaf3d: Fully convolutional anchor-free 3d object detection
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.CV 2representative citing papers
GraphFusion3D reports improved 3D object detection accuracy on SUN RGB-D and ScanNetV2 by combining adaptive image-to-point fusion with multi-scale graph reasoning on proposals.
citing papers explorer
-
PVCap: Towards Accurate 3D Dense Captioning via PseudoCap and VoxelCapNet
PVCap combines instance-mixing data augmentation with pseudo-labels and a voxel-based captioning network to achieve new state-of-the-art on 3D dense captioning benchmarks ScanRefer and Nr3D.
-
GraphFusion3D: Dynamic Graph Attention Convolution with Adaptive Cross-Modal Transformer for 3D Object Detection
GraphFusion3D reports improved 3D object detection accuracy on SUN RGB-D and ScanNetV2 by combining adaptive image-to-point fusion with multi-scale graph reasoning on proposals.