Pith. sign in

Paper Citation Record · LEDGER

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

As of 9 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 2 inbound Pith citation observations for arXiv:2506.06600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06600 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:57:22.751107Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T19:39:46.922877Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:24:01.588708Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy9
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c57bfb3f-f07d-4a98-a200-4525fd253797 · outbound

This paper cites Language Models are Few-Shot Learners.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Language Models are Few-Shot Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.627032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.627032Z digest=sha256:cbaccac7c03488851efe2a619585c46f06015e7147044698f46337d9ff22ca12

Observation 20461b7d-3f2e-4d9e-96ff-b7c3c5a2d8d5 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints LLaMA: Open and Efficient Foundation Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.632278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.632278Z digest=sha256:76d29784b904b0828690e6ebf3a559bddff685244f2f038bb3e77010c22ca0a4

Observation 21a63095-f6ba-463f-9676-4a004d5f9353 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.637181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.637181Z digest=sha256:24573a551aa336d8c6b888cffea03a1f2b873c6fdcd67418290815a5941e2545

Observation 05346d2e-3ece-4cfa-93fd-f83fcabbd77e · outbound

This paper cites The Llama 3 Herd of Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.641263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.641263Z digest=sha256:2c7cd4405fc099fb4ec95acba3111cb747d2857e449ceed857d505a2e8ed5b4f

Observation d355fdae-063a-4c0b-890d-c6dbc462590b · outbound

This paper cites Qwen3, April 2025.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Qwen3, April 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.193547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.646242Z digest=sha256:914289066f42b6778b2793b13ec2ca614b9e933b14c5af597e806380e31baf7f

Observation a27ca818-e6c9-4536-ba0d-7a5c6713cedd · outbound

This paper cites Visual-language models for medical image analysis: A survey.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Visual-language models for medical image analysis: A survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.182550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.650068Z digest=sha256:8c7df262b6f463f0e8fa25117923b5ce988b5c3ce5d5be425815eca549267c6e

Observation 3d146fbb-eb46-4603-916d-570fde5b3fc8 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Med-flamingo: a multimodal medical few-shot learner

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.654852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.654852Z digest=sha256:3cef2da8cb82e91606ccea329b4b70d0209de571c3f5db10d142023b10741b24

Observation 22fc5508-addf-460e-b57b-05dff12354d4 · outbound

This paper cites GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:57:22.982969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.658151Z digest=sha256:a0ffa89c7d80530c99bcbee5e5378e917cf7ceb6e9e214331ac5ae2c4235eaa5

Observation e4811ea2-b637-4362-a90f-f8fe9f73e1ad · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.662714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.662714Z digest=sha256:b3dad297616a9d5a4431086d7638a97f0cb532b667507f98e61bc41398e395c4

Observation b8a46e70-17ed-4b53-9668-28f92aec6133 · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.666851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.666851Z digest=sha256:98f47f795d5c319c60f304843a21f866be6cfdef0f0b6c6ae1a90d80c68df835

Observation 621c8906-edb3-44d1-83ca-f6fdc282c21e · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.670308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.670308Z digest=sha256:1182b331eb968ff51cc90c2408c757b5cf205a7b4396272c1a3fa708ae621002

Observation 48f93aba-269a-4d53-b74d-ce843f717b7f · outbound

This paper cites Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.673917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.673917Z digest=sha256:fea88df06699fc1478251b491b9ab824a7883ae98b288071fafaa9bea739842e

Observation 1aad558e-94e1-447e-bc96-998d294aecf9 · outbound

This paper cites Robust Kalman Filters Based on the Sub-Gaussian $\alpha$-stable Distribution.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Robust Kalman Filters Based on the Sub-Gaussian $\alpha$-stable Distribution

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.677585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.677585Z digest=sha256:9e67b8d1eb4b28894146e56d7c67b2a26f5be455a83b2e1e9cacbc3bfb2df5ee

Observation fca66905-e254-4a17-a9af-fc70bf6fc588 · outbound

This paper cites Artificial intelligence in radiology.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Artificial intelligence in radiology

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.142062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.681298Z digest=sha256:9a4a50dd66c257fbaed307ea25090c5c27802465632a089d3a799bd96a849eb5

Observation 6349b17c-70c1-4041-a3ed-97f290a22c00 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.684656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.684656Z digest=sha256:bfb608d237223b10ebbdff76baa7753ac302ccf1917e1777e267bd8cb7c00c3c

Observation eeb376da-144c-4f9e-bf4e-0ecab4f05ec3 · outbound

This paper cites Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.688075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.688075Z digest=sha256:e1d0e8b767e764876f8ad2e2a7080be236288b1db441d3261f2250962d224d5d

Observation d591a105-dcc3-4bbb-b57a-fa8afe9d2838 · outbound

This paper cites Explainability for artificial intelligence in healthcare: a multidisciplinary perspective.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Explainability for artificial intelligence in healthcare: a multidisciplinary perspective

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.130793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.691442Z digest=sha256:cc7feb4eb867205eafd3c970a3379d677ba039e40710e91112774f34be95618c

Observation dd1ecbe9-32cb-47c8-b2ae-cec7f1c9e7aa · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Chain-of-thought prompting elicits reasoning in large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.118859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.695048Z digest=sha256:d9ede5b2127aedd2759b3798afdcb259fed51dbe997a9d4838d1c310366f02dc

Observation c2ab2436-87e7-4f78-94c6-47d7493cd9ed · outbound

This paper cites MedCoT: Medical Chain of Thought via Hierarchical Expert.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedCoT: Medical Chain of Thought via Hierarchical Expert

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.698599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.698599Z digest=sha256:2ffed14655d61b97140ad13d23dd25ca129c86b991603bca4a8599a0460e0a4d

Observation b0fc92e6-b2f7-47af-b297-9e93e147bed8 · outbound

This paper cites Silvar-med: A speech-driven visual language model for explainable abnormality detection in medical imaging.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Silvar-med: A speech-driven visual language model for explainable abnormality detection in medical imaging

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.105655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.702267Z digest=sha256:b61e9c26f84ea4f5cbaf47178bf35f0b6c8698baa4240ab55653405e5292c407

Observation c1813c07-a182-4df5-af8e-fb5791917559 · outbound

This paper cites Two-Stage Estimation and Variance Modeling for Latency-Constrained Variational Quantum Algorithms.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Two-Stage Estimation and Variance Modeling for Latency-Constrained Variational Quantum Algorithms

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:57:22.845736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.706142Z digest=sha256:e308ba0a4acfc5aaa5c6f5f1e9f1a126580e708dfd8d879c024d2ec95bc0080b

Observation 0f22b2dd-101a-4481-9ea9-42dce54c8cf9 · outbound

This paper cites Reinforcement learning for medical image analysis: Current progress and future directions.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Reinforcement learning for medical image analysis: Current progress and future directions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.092508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.710524Z digest=sha256:919de34c7a7c5852c298251bb081c0454a6a9fcce33cc63e6faa1741773bd8d9

Observation fe64f8fd-0251-4290-99d3-5e70ee5715fa · outbound

This paper cites Which shapes can appear in a Curve Shortening Flow Singularity?.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Which shapes can appear in a Curve Shortening Flow Singularity?

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.713892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.713892Z digest=sha256:e2d21945f7af18da4d88bf6d413e1a296803e3ad11cde79499dae009cf501885

Observation a3f3f9be-363b-4006-9834-d829cd011408 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.717492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.717492Z digest=sha256:f3d95f9efec374cac7f6510320662bf8740115142ce612bd21ce5d8e3fd2a678

Observation 34f1c314-0099-4ed5-96ae-1d085cc868c4 · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints A dataset of clinically generated visual questions and answers about radiology images

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.721052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.721052Z digest=sha256:426f8bbbd7d5719d0f18ce4f15c51a4f95be13b9deeea3e741201a995ee308e6

Observation 0758ad61-c5fa-4d6f-b6c8-001ca2c3f319 · outbound

This paper cites SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.724573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.724573Z digest=sha256:ad4043e20e4964e24cc80ff0239200558f3679bd7ab976083cc208a2aca19561

Observation 8c7b02f2-f134-4049-86c5-81af31d62f94 · outbound

This paper cites Hasan, Vivek V.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Hasan, Vivek V

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.074082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.728235Z digest=sha256:27a8e67b44d8b22871482e217b3c98baf1c3580697ca70a9e632e8188164d047

Observation c5d2c434-a1a8-4ce4-aaa3-5290bd92de62 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.731580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.731580Z digest=sha256:a7d8a58188f31ce51f1464b5a90f6da75e02be4b5f196e8fc2b55e219f5ad25a

Observation 55493f64-ddfc-4e72-b236-94359fc45321 · outbound

This paper cites Weinberger, and Yoav Artzi.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Weinberger, and Yoav Artzi

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.063149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.735194Z digest=sha256:f26090c0d3f2c61171efe96a8331cc34378ac2dc0d65c953fcda649b2a77d6a5

Observation cedffaf2-33a3-4aaf-a605-84f17b598694 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Flashattention: Fast and memory-efficient exact attention with io-awareness

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.742808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.742808Z digest=sha256:21d6be1532a4fea0d20541d4c51ef8b577c320e54451f7a595c011ae3ece4421

Observation d107493a-58e3-434b-8b63-5de0e80d7d93 · outbound

This paper cites Cambrian-1: A fully open, vision-centric exploration of multimodal llms.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Cambrian-1: A fully open, vision-centric exploration of multimodal llms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.747458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.747458Z digest=sha256:68e67ee6422bd5829083999cb1e97fdd3a367b43394594c612ab87db3e249b1f

Observation a1a0afb0-d962-481f-9147-dd3b48a2456a · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.751107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.751107Z digest=sha256:1d58cbf32bb85c969b7de2da50dfb021bb7ab4407bb4145cc56d4d1c04a335b7

Observation 04006296-ec90-4526-8845-f0df9fa6fb01 · outbound

This paper cites an unresolved cited work.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Unresolved cited work

Reference 2020

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:57:23.051792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:57:22.738698Z digest=sha256:7e580fec97440ee325e98f1a77f632793b9c6185a7c90535aec5fc48f0b2a761

Pith citing papers

Observation a7916707-232e-4e32-99ff-9c13587ee46b · inbound

LLM-as-a-Judge in Healthcare: A Scoping Analysis of Applications, Methods, and Human Alignment cites this paper.

LLM-as-a-Judge in Healthcare: A Scoping Analysis of Applications, Methods, and Human Alignment RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:01.590061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T23:18:59.283834Z digest=sha256:fffb2a373239121bc653b2046b5d175e20e62d86ca5f6f019ed512d96bcf7131

Observation 03ad372f-6803-41f7-bc9d-8f8783da5a6d · inbound

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO cites this paper.

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T19:39:46.922877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:39:46.922877Z digest=sha256:3339a1bf7d869c5b8b0f2a789cc68dcbe9e3a21af0a84079efe8fad6d1c3a7c8