UniCT Depth fuses event and image data with a CNN-Transformer hybrid to improve monocular depth estimation, achieving lower average errors on MVSEC and DENSE benchmarks.
A 240 180 130 db 3 s latency global shutter spatiotemporal vision sensor
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
UniCT Depth: Event-Image Fusion Based Monocular Depth Estimation with Convolution-Compensated ViT Dual SA Block
UniCT Depth fuses event and image data with a CNN-Transformer hybrid to improve monocular depth estimation, achieving lower average errors on MVSEC and DENSE benchmarks.