CaST-Bench creates a benchmark with causal-chain annotations and novel metrics showing that current VLMs struggle to construct precise grounded causal chains in video QA.
En- hancing video-llm reasoning via agent-of-thoughts distilla- tion
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.CV 2years
2026 2representative citing papers
EFlow improves long-video QA by separating clip-finding from answering and adding a confidence trigger that re-reads the full video when unsure.
citing papers explorer
-
CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering
CaST-Bench creates a benchmark with causal-chain annotations and novel metrics showing that current VLMs struggle to construct precise grounded causal chains in video QA.
-
EFlow: Learning Evidence Flow for Long-Video Reasoning with Adaptive Reflection
EFlow improves long-video QA by separating clip-finding from answering and adding a confidence trigger that re-reads the full video when unsure.