LGA decomposes action labels and videos into three aligned atomic phases, then fuses text and video features to set a new state of the art in few-shot action recognition.
Depth guided adaptive meta-fusion network for few- shot video recognition
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
LGA decomposes action labels and videos into three aligned atomic phases, then fuses text and video features to set a new state of the art in few-shot action recognition.