Pith. sign in

Paper Citation Record · LEDGER

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs

As of 13 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2605.24492.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24492 v1

Coverage vector

measured 11 of 11 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T14:03:39.123887Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

11 of 11 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a48e7783-cc04-45f6-9bc4-267c56cd46d5 · outbound

This paper cites Qwen2.5-VL Technical Report.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Qwen2.5-VL Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T14:04:44.348547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:c3878d67d090cbb2c5bbd5a0be5ef131edf5468718f8abbefd8eac472989086d

Observation be4187e3-61d1-41e6-8bab-277e8ed0366d · outbound

This paper cites Improving Medical Diagnostics with Vision-Language Models: Convex Hull-Based Uncertainty Analysis.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Improving Medical Diagnostics with Vision-Language Models: Convex Hull-Based Uncertainty Analysis

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.351117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:ee81e269235b06ee2ad128661a0a62d199291d65cdac4cc28b1437d45bb493e5

Observation 9dfe3684-2199-44ef-981e-41c8ea065f9d · outbound

This paper cites GPT-4o System Card.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs GPT-4o System Card

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T14:04:44.345634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:25a3e5dbba8012cbce5736e28ea3583e809f0e605de30168680c27fc3ae21165

Observation f342570b-7945-4509-984e-ffffe6e21637 · outbound

This paper cites Multimedeval: A benchmark and a toolkit for evaluating medical vision-language models.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Multimedeval: A benchmark and a toolkit for evaluating medical vision-language models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T14:04:44.336551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:b40640e78dac82cbd2105c563baf68c3831abf196554accc3a07f06307dd7077

Observation 5fdc7609-8382-4ad9-bb0d-246ebd0ad317 · outbound

This paper cites Capabilities of Gemini Models in Medicine.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Capabilities of Gemini Models in Medicine

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T14:04:44.342060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:cf42a37d19ddba1b5d471111485dca57d5513972bc73d776281526d5417c5d45

Observation 0e38be63-c7b5-47bc-be97-804ce4099213 · outbound

This paper cites SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T14:04:44.350657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:e9cc98fe1255d0a0e66cf0bcba96cad037f530d042349f91b03697a84df40f9c

Observation 6c6facfa-4dfc-401d-bc81-b39660666d81 · outbound

This paper cites DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.343130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:4a778eeccde9b864d487934a9cf1524696887f583ff71e400938d38651158390

Observation 4fb0c62a-fc65-48d6-a0fc-945407c4c1e8 · outbound

This paper cites an unresolved cited work.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-07-08T23:55:43.959720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:5fec4b7ea063a579ae2b7612a7602ea04ce4d12e739e3be91688e89422fe5b7d

Observation 457497f2-3f41-40a7-9da1-4e2efe73d191 · outbound

This paper cites an unresolved cited work.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-07-08T23:55:43.958146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:bd996e72f0219af366d62656ae4db8866e5450e0375dc9dc211bf2e41f4ec283

Observation a3820aa5-7de8-4e3c-b809-790acdb10af0 · outbound

This paper cites an unresolved cited work.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-08T23:55:43.953930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:cf2fc476d514ff0486cba6a900a9b40d1e92d547de5be634c23a8492e8021d7e

Observation a160af29-9aec-4f1b-8c58-2b6dd3e64484 · outbound

This paper cites no lung opacity.

Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs no lung opacity

Reference 11

Resolution
malformed identifier
raw_fallback, observed 2026-07-08T23:55:43.955907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T14:03:39.123887Z digest=sha256:9c381289c35b7ce327d2df5e6d104e29e8ce56fd6509a82f36664e8fc8db28c9

Pith citing papers

No inbound Pith citation observations are available.