Pith. sign in

Quo vadis, action recognition? A new model and the kinetics dataset

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2024 1

verdicts

CONDITIONAL 1

representative citing papers

Efficient Transfer Learning for Video-language Foundation Models

cs.CV · 2024-11-18 · conditional · novelty 6.0

MSTA, a multi-modal spatio-temporal adapter with an LLM-based consistency constraint, achieves strong base-to-novel and few-shot video recognition performance using only a small fraction of trainable parameters.

citing papers explorer

Showing 1 of 1 citing paper.

  • Efficient Transfer Learning for Video-language Foundation Models cs.CV · 2024-11-18 · conditional · none · ref 1

    MSTA, a multi-modal spatio-temporal adapter with an LLM-based consistency constraint, achieves strong base-to-novel and few-shot video recognition performance using only a small fraction of trainable parameters.