Stereo-DEVO extends monocular deep event VO (DEVO) with a cheap static stereo association method, producing metric-scale, real-time pose estimates robust to HDR and aggressive motion.
Unsupervised Learning of Dense Optical Flow, Depth and Egomotion from Sparse Event Data
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
In this work we present a lightweight, unsupervised learning pipeline for \textit{dense} depth, optical flow and egomotion estimation from sparse event output of the Dynamic Vision Sensor (DVS). To tackle this low level vision task, we use a novel encoder-decoder neural network architecture - ECN. Our work is the first monocular pipeline that generates dense depth and optical flow from sparse event data only. The network works in self-supervised mode and has just 150k parameters. We evaluate our pipeline on the MVSEC self driving dataset and present results for depth, optical flow and and egomotion estimation. Due to the lightweight design, the inference part of the network runs at 250 FPS on a single GPU, making the pipeline ready for realtime robotics applications. Our experiments demonstrate significant improvements upon previous works that used deep learning on event data, as well as the ability of our pipeline to perform well during both day and night.
citation-role summary
citation-polarity summary
fields
cs.RO 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Deep Visual Odometry for Stereo Event Cameras
Stereo-DEVO extends monocular deep event VO (DEVO) with a cheap static stereo association method, producing metric-scale, real-time pose estimates robust to HDR and aggressive motion.