Pith. sign in

arXiv preprint arXiv:2504.01886 (2025)

3 Pith papers cite this work. Polarity classification is still indexing.

3 Pith papers citing it
abstract

Recent advances in general medical AI have made significant strides, but existing models often lack the reasoning capabilities needed for complex medical decision-making. This paper presents GMAI-VL-R1, a multimodal medical reasoning model enhanced by reinforcement learning (RL) to improve its reasoning abilities. Through iterative training, GMAI-VL-R1 optimizes decision-making, significantly boosting diagnostic accuracy and clinical support. We also develop a reasoning data synthesis method, generating step-by-step reasoning data via rejection sampling, which further enhances the model's generalization. Experimental results show that after RL training, GMAI-VL-R1 excels in tasks such as medical image diagnosis and visual question answering. While the model demonstrates basic memorization with supervised fine-tuning, RL is crucial for true generalization. Our work establishes new evaluation benchmarks and paves the way for future advancements in medical reasoning models. Code, data, and model will be released at \href{https://github.com/uni-medical/GMAI-VL-R1}{this link}.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 2 cs.AI 1

years

2026 3

roles

background 1

polarities

background 1

representative citing papers

citing papers explorer

Showing 3 of 3 citing papers.