SSTrack trains a Vision Transformer tracker without frame-wise box labels by combining forward global search, backward local association, and instance contrastive learning, and reports state-of-the-art self-supervised results on nine tracking benchmarks.
F.; Vedaldi, A.; and Torr, P
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Decoupled Spatio-Temporal Consistency Learning for Self-Supervised Tracking
SSTrack trains a Vision Transformer tracker without frame-wise box labels by combining forward global search, backward local association, and instance contrastive learning, and reports state-of-the-art self-supervised results on nine tracking benchmarks.