A visual player-identification module feeding names into an LLM improves identity-aware basketball video captioning; a new 9,726-clip dataset is released.
Quo vadis, action recognition? a new model and the kinetics dataset
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning
A visual player-identification module feeding names into an LLM improves identity-aware basketball video captioning; a new 9,726-clip dataset is released.