Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:34.049540Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.13973.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:34.049540Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c993af9c-0799-4b55-bc75-1b967876b1d8 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0086c278-2456-4593-9740-5be1e49a590d · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2625a23-56a5-45cf-900f-5be9c59664b9 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e585a3a3-fa6d-4b21-9c28-ec369a8b097a · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eca15f64-1a27-4119-821e-080a268450fd · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39147750-bcef-4a0f-885f-f886e0ed0be5 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-Language Models Can Self-Improve Reasoning via Reflection
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0353c2e0-0333-4069-944a-6ff0ea14d591 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c12872c-2522-40a6-970b-18415343905b · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 66edf0f7-809f-4720-b248-3123724d02d7 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a0aab34-6632-4de3-b56f-189c9c036697 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models The Llama 3 Herd of Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cdb0920-2151-4686-94cf-331907972674 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd43bf7-d09b-4d95-9382-1b4d39545788 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c09e7ae-55fe-4bb5-bc2d-714a569c4bb0 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b9ac47-5b2f-41f1-a935-572a3768597d · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Mistral 7B
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a7f238e-3dc0-49c4-a70a-ea85c494cf12 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models LLM Post-Training: A Deep Dive into Reasoning Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3224970-441f-48be-9339-575cec1021d1 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3db54ed8-b633-4a5d-8721-1bd9472187d9 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c59230b9-8a55-4f08-a5e8-635c00080c36 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2bbb2501-e7b7-4296-9881-50c504fd81a5 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff045b84-e5de-4954-a5aa-fab2900f7ee1 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Understanding R1-Zero-Like Training: A Critical Perspective
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c37a488c-14a6-4e37-930d-03174f64e814 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 85bc92cc-e12d-40b5-bf4b-85aacc88b35b · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10625c65-f303-48e7-9c90-4ced387d8e12 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc79f665-816d-45ab-98a2-834e7b962796 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a1e9c79-23c0-470c-9cb4-92c09170525a · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Proximal Policy Optimization Algorithms
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d97a5d0-1e19-4bfa-a1dc-496777c6c5cd · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b69e7c64-86ed-48bb-8779-7f4c0fb824f7 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0ffa0d3-726f-4542-871f-6a653604d65a · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e21db740-c644-4fd7-bdcd-e6fd90737047 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1947b600-f53c-449e-843c-2386d4acd9ef · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a6b834a-b81e-450e-8489-5efb40855b7e · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9e3019e2-026c-4072-bb84-e23a7b886cd7 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06a2df12-52da-4aee-95e4-e9c4a025b307 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Improve Vision Language Model Chain-of-thought Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45eb7a77-1544-45be-9be2-8045541f163e · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a8f6d7-957d-4702-89d6-8ebcfda7b726 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46c66b0-777a-4121-bda9-b6f6f8947210 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4591e73e-4ce8-40f9-ae02-e5d880d8d484 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 52a61e03-e4ca-4a96-86a6-fee5261db3a2 · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 247b3255-68d8-431b-8ceb-1ecf28d49abd · outbound
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models RetinalGPT: A Retinal Clinical Preference Conversational Assistant Powered by Large Vision-Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.