A challenge report showing that MLLM-generated scene graphs and ConceptNet knowledge graphs each give small accuracy gains over a video-only baseline, and that per-category selection reaches 44.21% on the HD-EPIC VQA benchmark.
Comet: Com- monsense transformers for automatic knowledge graph con- struction
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
method 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
method 1polarities
use method 1representative citing papers
citing papers explorer
-
From Pixels to Graphs: using Scene and Knowledge Graphs for HD-EPIC VQA Challenge
A challenge report showing that MLLM-generated scene graphs and ConceptNet knowledge graphs each give small accuracy gains over a video-only baseline, and that per-category selection reaches 44.21% on the HD-EPIC VQA benchmark.