DHNet with patch alignment and dual hypergraph fusion reaches SOTA RGBT video object detection on VT-VOD50 and the new large-scale DVT-VOD1000 benchmark.
Icafusion: Iterative cross-attention guided feature fusion for multispectral object detection,
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.CV 2years
2026 2representative citing papers
DSAFormer applies spatial and channel sparse multi-head cross-attention plus a learnable fusion block to reduce redundant token interactions and improve multispectral object detection on four public datasets.
citing papers explorer
-
Dual-Correlation Hypergraph Network for Unaligned RGBT Video Object Detection and A Large-scale Benchmark
DHNet with patch alignment and dual hypergraph fusion reaches SOTA RGBT video object detection on VT-VOD50 and the new large-scale DVT-VOD1000 benchmark.
-
Dual Sparse Aggregation Transformer for Multispectral Object Detection
DSAFormer applies spatial and channel sparse multi-head cross-attention plus a learnable fusion block to reduce redundant token interactions and improve multispectral object detection on four public datasets.