F-SUM, an image-computable score combining foveated vision and vision-language models, correlates with human scene comprehension times (r=0.47) and saccade counts (r=0.51) across 277 images.
Neurology research international 2014, 301473
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps
F-SUM, an image-computable score combining foveated vision and vision-language models, correlates with human scene comprehension times (r=0.47) and saccade counts (r=0.51) across 277 images.