Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:24:42.605024Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 6 inbound Pith citation observations for arXiv:2505.19213.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:24:42.605024Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:00:49.780513Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T10:15:45.170778Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 083e3c0e-d94f-4e1d-9da8-01239ac2f2e8 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 485406a0-4f21-4692-9d72-13746cd7c810 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Curriculum learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4984e00a-9b04-4f94-9a48-9fc79a1f61ca · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4f7ae01-793e-4ea3-90f5-516f7ba1e176 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 414891cd-b1a4-4552-9e48-849577895210 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d957eeb-e9f8-477f-8c4b-6e156613f329 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 194b85c6-e245-417c-be55-fc959ed2dcad · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning PathVQA: 30000+ Questions for Medical Visual Question Answering
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc7f90b0-980f-4fb5-8fb5-3d6f445c2311 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Omnimed- vqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7294fd7b-2b50-4a7d-8e47-73c13e29ea10 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f572d3-4d46-4a38-97da-8872d8fef298 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Gonzalez, Hao Zhang, and Ion Stoica
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16a6aac7-759d-4c74-9caf-8ada360f6976 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f3ec62-c8eb-49c3-b21a-2d205f4ca00f · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning A dataset of clinically generated visual questions and answers about radiology images.Scientific data, 5(1):1–10, 2018
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 924a7d54-5ccd-42f7-a856-9053fd5aeba6 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64b2524f-111e-48c3-88b4-45eb68fc228b · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Llava-med: Training a large language-and-vision assistant for biomedicine in one day.Advances in Neural Information Processing Systems, 36:28541–28564, 2023
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cb2d57b-458f-46c5-a3f0-de02b9b13686 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c2e270-d74b-4a94-9f9d-1f2ff245b5f5 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Slake: A semantically- labeled knowledge-enhanced dataset for medical visual question answering
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 743d4455-6f6c-46ed-9c84-51601b7f11b7 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e165dee1-00b5-4151-89af-8d772649130d · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30310768-6994-47ff-8bc7-445dbc9c98a9 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04e546ed-1663-4d8a-a641-f624cd64382b · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Med-flamingo: a multimodal medical few-shot learner
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16f03c70-230e-4852-847a-cf1b281a323c · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c757ce95-577b-468f-a5aa-a7f73106c3ed · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Learning transferable visual models from natural language supervision
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3229f476-5b6a-4b7f-9d02-e3d08ae3bcaf · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9de819ef-1b09-46d3-95c6-3a74402fe734 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Quilt-llava: Visual instruction tuning by extracting localized narratives from open- source histopathology videos
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9508b1-cedb-452b-9d13-dbb420cc2cdf · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38209900-4f7e-46df-a766-5a5b326affb1 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dfb0a38-c2b7-4582-9303-d6f0c64ab1d6 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01b412b5-8ee6-4f45-b4e9-76c18a49d7c9 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f66319b4-9117-472e-9c0e-83cf91026761 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Qwq-32b: Embracing the power of reinforcement learning, March 2025
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c994559-1c4f-4a4e-a084-4f9d5dee187f · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning VisualPRM: An Effective Process Reward Model for Multimodal Reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f87a5b-8251-4fa6-865c-5da7a6681925 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Chain-of-thought prompting elicits reasoning in large language models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f426e4a-38af-491c-85cf-cda98f68442e · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a465c906-8a80-426f-be98-5d5d5a5b2b6a · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63a3a088-1330-4acf-87fb-41c8a6bf9bc7 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Yi: Open Foundation Models by 01.AI
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aa50367-b97b-4643-ab22-b3351f0a2e56 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c688696-33d5-4283-99ed-7a871a8cb5f4 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62d3e441-4423-4389-b657-bff9f78b355e · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Rlaif-v: Aligning mllms through open-source ai feedback for super gpt-4v trustworthiness.arXiv preprint arXiv:2405.17220, 2024
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cf5abd0-9e3c-46ce-a134-fd05dc177399 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59ac7428-3669-4f1a-8ee1-344b0f87e5e7 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c884f0-3a42-4aea-af1e-14ff3c327dce · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fa16a12-2e15-4336-8e07-ae54486d285e · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc64ed3d-8be0-4113-b7ad-b4f368c8077b · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Llamafactory: Unified efficient fine-tuning of 100+ language models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d79c18-f00e-4afb-b695-829c38aed203 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd59a4c8-da6f-43b8-9998-f4168e7382ee · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning Aligning Modalities in Vision Large Language Models via Preference Fine-tuning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 247bc90c-215b-453a-ad87-185f4f606d53 · outbound
Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67b2389a-d06d-445b-a313-a7fe8f0411be · inbound
CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bd09e0e-b0e5-45bf-b46e-894236e74223 · inbound
AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3beaf0a-ec85-472b-8a70-3919146c80c0 · inbound
A global log for medical AI Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 117
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f11b0e0-75db-4281-8571-257bbcc4e6bb · inbound
Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c9da16b-edaf-4586-9088-323df898fe22 · inbound
Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f4a817e-1ce2-40c8-a068-9739e97f1dd7 · inbound
Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.