Pith. sign in

Paper Citation Record · LEDGER

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.13973.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.13973 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:34.049540Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c993af9c-0799-4b55-bc75-1b967876b1d8 · outbound

This paper cites online" 'onlinestring :=.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:26.837843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:26.837843Z digest=sha256:658a66bfe3d4b941b569b120092b34d29ef6abcd1ccff3e7ffba3d6c404d9950

Observation 0086c278-2456-4593-9740-5be1e49a590d · outbound

This paper cites write newline.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.116351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.116351Z digest=sha256:3486c3b3c906db37baa521bc0249501b065726eccc171bd5d040f8f6dc524bf4

Observation d2625a23-56a5-45cf-900f-5be9c59664b9 · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.245226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.245226Z digest=sha256:67b607e828dc07312575f4f981000a502128df94a33910781d2bcc673c733d75

Observation e585a3a3-fa6d-4b21-9c28-ec369a8b097a · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.830709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:30.351330Z digest=sha256:c74207dd280defa145dc69aca2251fa0ad9f608a19487e03d193b89d5caf03b3

Observation eca15f64-1a27-4119-821e-080a268450fd · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.454995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.454995Z digest=sha256:0e853a7ab48129923d8de120d9b0cd6a3ee4e4301a7bd39d3e3ef3d343da139f

Observation 39147750-bcef-4a0f-885f-f886e0ed0be5 · outbound

This paper cites Vision-Language Models Can Self-Improve Reasoning via Reflection.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-Language Models Can Self-Improve Reasoning via Reflection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.548612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.548612Z digest=sha256:4c1e1d3d570718a3759cec7f7b8d7c6506bc70ba0d073f0b21c52b511d88aac6

Observation 0353c2e0-0333-4069-944a-6ff0ea14d591 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.665481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.665481Z digest=sha256:85ef0ae79c786229a57fb99774d9d674bf8aac05f51cab2462fa178991f832f0

Observation 7c12872c-2522-40a6-970b-18415343905b · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.614745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:30.752803Z digest=sha256:105356fdb5d79f36c320a8d38ef8a79c517411c428f703c40f9a1ca10ceda2b5

Observation 66edf0f7-809f-4720-b248-3123724d02d7 · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.841430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.841430Z digest=sha256:e7314384f55634659bf3baf6056340361c5c43bf3da4a66b3f2f0cb8d7c2aa74

Observation 2a0aab34-6632-4de3-b56f-189c9c036697 · outbound

This paper cites The Llama 3 Herd of Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.931668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.931668Z digest=sha256:75ffb7bd0ea148e595535b8a64f63d3a4af9acca1e52b6c99adcf615410a88ca

Observation 4cdb0920-2151-4686-94cf-331907972674 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.075196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.075196Z digest=sha256:7d7d650801971730c86ed885e8f66eab48f098cef7753da73399065f37cd1287

Observation ccd43bf7-d09b-4d95-9382-1b4d39545788 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.214749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.214749Z digest=sha256:65c332479f4555d1519fac1ea80b5af9194401399ba6b8c0daa9d9f835b912f0

Observation 2c09e7ae-55fe-4bb5-bc2d-714a569c4bb0 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.317058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.317058Z digest=sha256:04046e324916456b142b6813e3fe78031c0b9bd1a089e98e1cec0d5f00fd4b31

Observation 98b9ac47-5b2f-41f1-a935-572a3768597d · outbound

This paper cites Mistral 7B.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.423898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.423898Z digest=sha256:b336d860a0667a5a0945e8feb30d336e2ee6db8e0308944cf4cd2455c809971d

Observation 9a7f238e-3dc0-49c4-a70a-ea85c494cf12 · outbound

This paper cites LLM Post-Training: A Deep Dive into Reasoning Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models LLM Post-Training: A Deep Dive into Reasoning Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.517756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.517756Z digest=sha256:6420f34125a487c4e16362253467555c8b5b4b23fcd4962fa7642251138503e1

Observation d3224970-441f-48be-9339-575cec1021d1 · outbound

This paper cites BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.609026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.609026Z digest=sha256:83fc1684807923aad45207047ee5e68306509b654207da873eeaad18ba461b73

Observation 3db54ed8-b633-4a5d-8721-1bd9472187d9 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.706196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.706196Z digest=sha256:583e660a4c132758624b7b982ac444f063000cac766016bec7d1090a4d68f9a1

Observation c59230b9-8a55-4f08-a5e8-635c00080c36 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.372647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:31.790430Z digest=sha256:0a6fa43a36151bc3072ff78aa8afc554ad2ce031a4396a0ac74d3751e7dad68f

Observation 2bbb2501-e7b7-4296-9881-50c504fd81a5 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.871537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.871537Z digest=sha256:f7c123e2a0443372fd3f9a6d254c53ebc5dfd84ca4dad1a3faab0e7ef6096cb6

Observation ff045b84-e5de-4954-a5aa-fab2900f7ee1 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Understanding R1-Zero-Like Training: A Critical Perspective

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.932837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.932837Z digest=sha256:45f3eb7d980fc855bef7f9c40e44eeef5fd6071a7e12d1385d026b488ab0a960

Observation c37a488c-14a6-4e37-930d-03174f64e814 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.154763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.056482Z digest=sha256:db5fec8b9c47b2461de6b1cc82c1bf50000214b4e7b0c38b376fb87f254ab512

Observation 85bc92cc-e12d-40b5-bf4b-85aacc88b35b · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.164373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.164373Z digest=sha256:59672a92c566228f726a9a8069326b0c9aa7793041178e59c6fb45cf0e761895

Observation 10625c65-f303-48e7-9c90-4ced387d8e12 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.291467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.291467Z digest=sha256:37b70fa8498b5ee9a194f9e39f032562f3c6b955ddaec54637a5542d03381a07

Observation cc79f665-816d-45ab-98a2-834e7b962796 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.391411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.391411Z digest=sha256:b8874e9a2bfc4b9382ecc41e00a7de38c538370a233634e08f515fbfea315a28

Observation 8a1e9c79-23c0-470c-9cb4-92c09170525a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Proximal Policy Optimization Algorithms

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.491701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.491701Z digest=sha256:2d0dd236ca6ff5393fa117705aa050c56856511fe302ce28c9d4b7ac2f612b34

Observation 2d97a5d0-1e19-4bfa-a1dc-496777c6c5cd · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.625665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.625665Z digest=sha256:a68dac6124ee4533da003ebd4aaca60f5e5cb481aa48d5e93c7e533f682dfee3

Observation b69e7c64-86ed-48bb-8779-7f4c0fb824f7 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.721672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.721672Z digest=sha256:08e056577ea7f5665ecdd695ad57663bcbc8ab29d50e925e2d3518d06c23e039

Observation e0ffa0d3-726f-4542-871f-6a653604d65a · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.968000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.849962Z digest=sha256:938a487f6d83b20ffd368c72f1c9050ae72d60ecc86bac9e39f9c88d4e9a8cd3

Observation e21db740-c644-4fd7-bdcd-e6fd90737047 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.001856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.001856Z digest=sha256:7e41ef0263421cc455a4b87c5f0121f8bf9e5da0700668a9a8d4ec3fdfd888aa

Observation 1947b600-f53c-449e-843c-2386d4acd9ef · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.082866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.082866Z digest=sha256:af1ba6b2897f3efd2412de219890d2786366c58242ad6e3aad75f6bd9f744c99

Observation 9a6b834a-b81e-450e-8489-5efb40855b7e · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.834256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.221448Z digest=sha256:8cb25bd42fb94adaa11d613428a482e863d23932d6ecaa0fe3f2cb9880a604b5

Observation 9e3019e2-026c-4072-bb84-e23a7b886cd7 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.341166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.341166Z digest=sha256:3c0cab05aa573b92fc99eecd286b83e22cd62a81f7522001b4c2cbbaf8c31628

Observation 06a2df12-52da-4aee-95e4-e9c4a025b307 · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Improve Vision Language Model Chain-of-thought Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.432167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.432167Z digest=sha256:1591a2437c676eb9c7aa667f707ca5b0b28632504993209f7d4f22830d90cebb

Observation 45eb7a77-1544-45be-9be2-8045541f163e · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.540806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.540806Z digest=sha256:1b59b4a8de3cc62ffe3d1090b3d6f1a329f70d762a69aaf6d0669d45ca0f14f7

Observation 66a8f6d7-957d-4702-89d6-8ebcfda7b726 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.648501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.648501Z digest=sha256:9f11ec0197c1e8bfb674c0fd5b6c34fe7bde895562ee7ae2e9408c47063324e9

Observation e46c66b0-777a-4121-bda9-b6f6f8947210 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.729004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.729004Z digest=sha256:fce1a1bb661aee2d86bf1dac7d0e1cd7871bad4186a54eddcae5de71591d5932

Observation 4591e73e-4ce8-40f9-ae02-e5d880d8d484 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.611405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.787424Z digest=sha256:2c3af81cd7ec3dbd8a99670c58447471c5c668ccb5d650e2dc2a40328e7a9bce

Observation 52a61e03-e4ca-4a96-86a6-fee5261db3a2 · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.968558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.968558Z digest=sha256:ed398a4349d88c95f5023a72eb48bf749532595aae2539186ec4207397239294

Observation 247b3255-68d8-431b-8ceb-1ecf28d49abd · outbound

This paper cites RetinalGPT: A Retinal Clinical Preference Conversational Assistant Powered by Large Vision-Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models RetinalGPT: A Retinal Clinical Preference Conversational Assistant Powered by Large Vision-Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:34.049540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:34.049540Z digest=sha256:92645ee4690c61d7039a98cb947c55d4e5a75c2b4771eeef0be2a7f52d72e3f0

Pith citing papers

No inbound Pith citation observations are available.