Pith. sign in

Paper Citation Record · LEDGER

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

As of 14 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 2 inbound Pith citation observations for arXiv:2507.11200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.11200 v2

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:19:47.183398Z

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T13:30:26.488826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:59.467672Z

Reference resolution

8 of 8 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 740feadf-96d1-44be-8a17-9966a2751240 · outbound

This paper cites , " * write output.state after.block = add.period write.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study , " * write output.state after.block = add.period write

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.780825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.780825Z digest=sha256:1ecc64eea19f57b1a8a0af39be474ce4df69774889f4ec343705a5cd329d9e6e

Observation 1d22e2f0-6239-43fb-b5d2-e812226b0145 · outbound

This paper cites write newline.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.808585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.808585Z digest=sha256:6ce1a87894ccf684ac524349f53a5033b1ae6834d00e108dbeb0f3b399225358

Observation 2364e215-df6b-4c97-ad7a-176ea07231e5 · outbound

This paper cites Qwen2.5-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.853396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.853396Z digest=sha256:0ef3037a70a36bcc7352bf9e0d57b9513cbee3d81ee091f563793757d9e882e2

Observation faf3aa55-f2f1-43e7-b1d3-3c55f7592289 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.890585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.890585Z digest=sha256:d601972aa0fc347535dde0f5f8f101d535037497a796dc7bf445bfd2ebb23614

Observation 9c0eb45d-968e-4b77-abab-ae40351bd9f6 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.982626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.982626Z digest=sha256:78aa7f38f12711389c5365a37a082c0f6d31592c92ccc1c29c7e607ceeb7cbed

Observation 46e0d4a7-b639-4306-8547-51f65d2169d2 · outbound

This paper cites MiMo-VL Technical Report.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MiMo-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.032597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.032597Z digest=sha256:d03ee4a3f536ebdebcff62ea24453c977ecd1f972a69cb8038cf2e43e004f0c1

Observation a2fd42b1-9267-40cb-a7c9-dcdd005dd195 · outbound

This paper cites Disentangling Reasoning and Knowledge in Medical Large Language Models.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.083790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.083790Z digest=sha256:0b0d7aff18fe046f70be1b4827d58c930ce45c2ebcfc4436e8a4a9b6ed9fd24e

Observation 369e548c-b35f-4172-a194-0dc66d14b44b · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.183398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.183398Z digest=sha256:e1f8ad0b7f73509552d4a29b8988abef21d1e8c2ed543189f91590335652297d

Pith citing papers

Observation a4d1559b-01ff-4c57-80ba-495f6bd1e003 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:18:59.469375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T00:42:29.542446Z digest=sha256:9fa056bbeaed34a2af05ae802635699e3ed1d22b5a89d3c9575d82b187881896

Observation 0e712b36-152f-4be7-bb8a-d510af778c50 · inbound

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images cites this paper.

Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T13:30:26.488826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:30:26.488826Z digest=sha256:80462e0e5b4c07be34f422ece64671b041802e984f4555c8be1966fd69267705