A 4-billion-parameter vision-language model with a shared 3D MRI encoder and 4D rotary position encoding reports BERTScore 0.856 for mpMRI report generation and 0.912 multiple-choice accuracy on its internal dataset.
arXiv , year=
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Mr3D-VL: A generalist vision language foundation model for Multiparametric 3D Magnetic Resonance Imaging
A 4-billion-parameter vision-language model with a shared 3D MRI encoder and 4D rotary position encoding reports BERTScore 0.856 for mpMRI report generation and 0.912 multiple-choice accuracy on its internal dataset.