Pith. sign in

Paper Citation Record · LEDGER

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

As of 8 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 2 inbound Pith citation observations for arXiv:2507.11200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.11200 v2

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:19:47.183398Z

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T13:30:26.488826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:59.467672Z

Reference resolution

8 of 8 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 740feadf-96d1-44be-8a17-9966a2751240 · outbound

This paper cites , " * write output.state after.block = add.period write.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study , " * write output.state after.block = add.period write

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.780825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.780825Z digest=sha256:3c76b1bb497cfbe45e65aa03a5ab56a2262e47e2e64760d3f7a4e1888663d64d

Observation 1d22e2f0-6239-43fb-b5d2-e812226b0145 · outbound

This paper cites write newline.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.808585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.808585Z digest=sha256:7db581df650224a4ecbf25318f4bc8b02fb3762beadb05754bc18556381a4da5

Observation 2364e215-df6b-4c97-ad7a-176ea07231e5 · outbound

This paper cites Qwen2.5-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.853396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.853396Z digest=sha256:f97a3363a65c82db91fe14ada30fba7dfd0f22cf11219be82b6f38fbfa53b1ac

Observation faf3aa55-f2f1-43e7-b1d3-3c55f7592289 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.890585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.890585Z digest=sha256:e2e7af88c3d6f49a45f613a2ddea3aee30204574fdb1fb584dc0bb6fbb90274a

Observation 9c0eb45d-968e-4b77-abab-ae40351bd9f6 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.982626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.982626Z digest=sha256:45f3f1b45c8ed670e4bd85d25779e94d0bde1c1079b2f529cd5912bfff393d19

Observation 46e0d4a7-b639-4306-8547-51f65d2169d2 · outbound

This paper cites MiMo-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MiMo-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.032597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.032597Z digest=sha256:54bcbaaf8fc1f3a847941ccbf8866614bca37adb69bf55fe06252b2d149b58be

Observation a2fd42b1-9267-40cb-a7c9-dcdd005dd195 · outbound

This paper cites Disentangling Reasoning and Knowledge in Medical Large Language Models.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.083790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.083790Z digest=sha256:12f88244f7b60216704ab04f726aacd1b5f4020b38d17f9d6a05155759dc402a

Observation 369e548c-b35f-4172-a194-0dc66d14b44b · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.183398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.183398Z digest=sha256:6eda8dab73b7041ceb669cbff3fbd57147876580f8aee1138348f9133d673547

Pith citing papers

Observation a4d1559b-01ff-4c57-80ba-495f6bd1e003 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:18:59.469375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T00:42:29.542446Z digest=sha256:60deef868f5adabab2b0829701e12a573dd0e1a66f00c99b944e367e4b0c97fa

Observation 0e712b36-152f-4be7-bb8a-d510af778c50 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T13:30:26.488826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:30:26.488826Z digest=sha256:95b24259d4906c669d8bfb21439c5603cff152105960eca763af722dab5d19ef