A 4D scene graph augmented with atomic human-object interactions and goal-driven events improves retrospective question answering about human activities in dynamic scenes.
A V A: A video dataset of spatio-temporally localized atomic visual actions,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes
A 4D scene graph augmented with atomic human-object interactions and goal-driven events improves retrospective question answering about human activities in dynamic scenes.