On the GEN1 event dataset, a Vision Transformer is reported to be more robust than a ResNet to simulated event noise, while the ResNet is slightly more accurate on clean data.
Deep residual learning for image recognition,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
From Ground to Air: Noise Robustness in Vision Transformers and CNNs for Event-Based Vehicle Classification with Potential UAV Applications
On the GEN1 event dataset, a Vision Transformer is reported to be more robust than a ResNet to simulated event noise, while the ResNet is slightly more accurate on clean data.