EyeMVP learns OCT-informed CFP representations via cross-modal masked reconstruction on 674k paired triples and reports competitive or superior performance on 15 retinal classification and segmentation tasks.
Masked au- toencoders are scalable vision learners
3 Pith papers cite this work. Polarity classification is still indexing.
years
2026 3representative citing papers
A frozen ImageNet MAE plus a per-dataset normalizing flow detects time series anomalies with average AUC-ROC 0.852 over nine datasets, the best average among 15 compared baselines.
SpO2-predictor-guided masked time-frequency PPG reconstruction yields subject-level MAE of 2.882% (OpenOximetry) and 2.359% (private wearable) versus stronger baselines.
citing papers explorer
-
EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining
EyeMVP learns OCT-informed CFP representations via cross-modal masked reconstruction on 674k paired triples and reports competitive or superior performance on 15 retinal classification and segmentation tasks.
-
VAN-AD: Visual Masked Autoencoder with Normalizing Flow For Time Series Anomaly Detection
A frozen ImageNet MAE plus a per-dataset normalizing flow detects time series anomalies with average AUC-ROC 0.852 over nine datasets, the best average among 15 compared baselines.
-
SpO$_2$ Predictor-Guided Stage-Wise Time-Frequency Reconstruction of Low-Quality Dual-Wavelength PPG for Oxygen Saturation Estimation
SpO2-predictor-guided masked time-frequency PPG reconstruction yields subject-level MAE of 2.882% (OpenOximetry) and 2.359% (private wearable) versus stronger baselines.