MindShot reconstructs multiple video shots from fMRI by first predicting shot boundaries, then using LLM-generated captions of keyframes to decode visual content.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
MindShot: Multi-Shot Video Reconstruction from fMRI with LLM Decoding
MindShot reconstructs multiple video shots from fMRI by first predicting shot boundaries, then using LLM-generated captions of keyframes to decode visual content.