Pith. sign in

Is space-time attention all you need for video understanding? In ICML, page 4, 2021

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

AdaVid: Adaptive Video-Language Pretraining

cs.CV · 2025-04-16 · conditional · novelty 5.0

AdaVid trains video-language encoders whose hidden dimensions can be stripped down at inference time, matching a standard model at half the FLOPs on EgoMCQ.

citing papers explorer

Showing 1 of 1 citing paper.

  • AdaVid: Adaptive Video-Language Pretraining cs.CV · 2025-04-16 · conditional · none · ref 5

    AdaVid trains video-language encoders whose hidden dimensions can be stripped down at inference time, matching a standard model at half the FLOPs on EgoMCQ.