KeepFIT V2 pretrains a fundus vision-language model with a small 'elite' image-text dataset plus public categorical labels, reaching performance competitive with models trained on much larger private data.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
MM-Retinal V2: Transfer an Elite Knowledge Spark into Fundus Vision-Language Pretraining
KeepFIT V2 pretrains a fundus vision-language model with a small 'elite' image-text dataset plus public categorical labels, reaching performance competitive with models trained on much larger private data.