An attention-based spatial-temporal feature blending extension to object detectors achieves 91.17% mAP at great ape detection on 500 camera trap videos, outperforming frame-based baselines.
Quo Vadis, action recognition? A new model and the kinetics dataset
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Great Ape Detection in Challenging Jungle Camera Trap Footage via Attention-Based Spatial and Temporal Feature Blending
An attention-based spatial-temporal feature blending extension to object detectors achieves 91.17% mAP at great ape detection on 500 camera trap videos, outperforming frame-based baselines.