Robust-U1 equips MLLMs with self-recovery via supervised fine-tuning, RL using SSIM and CLIP rewards, and joint multimodal reasoning, yielding SOTA robustness on corruption benchmarks.
When mllms meet compression distortion: A coding paradigm tailored to mllms.arXiv preprint arXiv:2509.24258, 2025a
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?
Robust-U1 equips MLLMs with self-recovery via supervised fine-tuning, RL using SSIM and CLIP rewards, and joint multimodal reasoning, yielding SOTA robustness on corruption benchmarks.