Pith. sign in

Paper Citation Record · LEDGER

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

As of 8 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2505.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15425 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:21.068140Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:13:53.947203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f9830d4a-77cd-49dc-bb44-3a02a7bea89d · outbound

This paper cites https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.441969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:17.569003Z digest=sha256:2752a27bb64fb4b7a4a16531f533d4ed242071f4cb2939c7e7cc201c9080dc6f

Observation ba554940-a11b-4daa-adf3-e009f5bc99d8 · outbound

This paper cites A Survey of Medical Vision-and-Language Applications and Their Techniques.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? A Survey of Medical Vision-and-Language Applications and Their Techniques

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.662259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.662259Z digest=sha256:1f544e15878a72bc251ff242aa1bbce09cd0cbf9e06d4ce116bb6b20f2f51bcc

Observation bf83a1f4-5215-408a-b498-db43f58fb80e · outbound

This paper cites In: ICML Workshop on Computational Biology (2021) 2, 9, 20.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML Workshop on Computational Biology (2021) 2, 9, 20

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.260674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:17.808789Z digest=sha256:4dd58b84ca33d044de444adf3b509f7c1bfe586fc355939039ceda980ca33a3e

Observation 915abbaa-6cd8-407d-94a5-815b401de255 · outbound

This paper cites Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.850405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:17.873767Z digest=sha256:784f94b92953150354aff90fc0a77959810ead941247713f6f7d968079ac6e1a

Observation 33df9c7a-f727-447c-b8b0-3f4c76bdfce1 · outbound

This paper cites MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.967185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.967185Z digest=sha256:da7d8373b337fb2c055ce95823faa9faf99fadcc9edd879278c4c3b84a429633

Observation 43739308-acd1-498c-aec4-5ee90fcd9201 · outbound

This paper cites In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.046793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.102723Z digest=sha256:91e77e36a72a15c7739d36ffaf11397a8d5a79bbda68d9497c307f2012cbaf4a

Observation 90a6e254-9ef8-4767-a418-54fd3f10fecd · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.860908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.204355Z digest=sha256:e9375f621ba8a04805a785cd586848b641a8b5ab58442c4e56985197db3e7292

Observation 956ac906-776d-4e26-aaf0-f4fd2f77d46c · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.653457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.321901Z digest=sha256:c97a68ed77665240e2a4ccd82816c1857077cb17981f521f8c5dbc8b58c4ec62

Observation c53fdf87-3b86-48f8-91ef-d0c64306771a · outbound

This paper cites MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.675948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.499866Z digest=sha256:391010c729c09dd79c0662c47e22dc032b169a5e99ab04b9963317ade9b916f3

Observation 58959005-40ed-403b-83fa-087010bf6644 · outbound

This paper cites In: ICLR (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICLR (2019) 2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.434967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.635490Z digest=sha256:61e1c72af4aee264a794007e1949c1ff49bad707159d39f53794919cc200e873

Observation c6332889-c3c1-4d21-a701-e48ba1998e42 · outbound

This paper cites ICLR1(2), 3 (2022) 8.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? ICLR1(2), 3 (2022) 8

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.253117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:18.776022Z digest=sha256:2a4bb50dfc1b3559d0aad5041af61cd5aac165fa6027796ecf89b5fa1ce7db0b

Observation 4ef916d3-dec1-4414-b183-18395aa5422f · outbound

This paper cites Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:18.969671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:18.969671Z digest=sha256:78071b9ed4a92ee7cdeb6ab20e2b746e3a4812552b25d272274775536534e8c0

Observation bd4549b0-3e2a-421f-bf14-7c5a85841c67 · outbound

This paper cites Noise is an Efficient Learner for Zero-Shot Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Noise is an Efficient Learner for Zero-Shot Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.132444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.132444Z digest=sha256:94e2941a5976f977b51368c7aa4b2687a0cc476d10756658fcbc3fa214cddf30

Observation f675dad8-f92c-4e60-b223-63bef8bb45ef · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.961842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:19.292556Z digest=sha256:48d29846ef4ec2f96d1941370631e221b2f089a6a55d26a69993c1c78cab66d9

Observation dc13bbbe-e59a-4483-92ac-88d0c26f5846 · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.434663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.434663Z digest=sha256:4e6108ed891213622e8d5e6144ae6aeef8fe5bccc0b780bc4c664d4203e697f0

Observation 1ef37400-0521-4c12-a2e3-0e6c5992fbd1 · outbound

This paper cites Radiology p.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Radiology p

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.735407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:19.547418Z digest=sha256:213624df3f1135ede9127dc2a64dae58319105e7ba48c8073afc3614879d950a

Observation add91f25-6bb1-405f-bc53-8dc00554c113 · outbound

This paper cites IEEE Reviews in Biomedical Engineering (2025) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? IEEE Reviews in Biomedical Engineering (2025) 2

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.484419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:19.662098Z digest=sha256:3045214e725fb3e8103ecfd9835d826be98894d31b45c1ae153aea5d46cb6ffb

Observation 85331b14-0378-4f8b-b6ba-70500ef184b5 · outbound

This paper cites UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.892988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.892988Z digest=sha256:169fcc1472d007f8c13aa2538a93e195b12280e6a68ed9b9398532e55947eede

Observation fcfb872b-e090-4485-911c-a27521c16616 · outbound

This paper cites Journal of Biomedical Informatics135, 104234 (2022) 4.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Journal of Biomedical Informatics135, 104234 (2022) 4

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:19.989329Z digest=sha256:a5730f9b5ffef17ed5bea25dc37abf59e5738dae9d890b5445ab0177c784611e

Observation d36f2d04-e14a-40b2-96e8-927b071b9a66 · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.936325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.112534Z digest=sha256:8de168d5a30aa143b791343631acfc590511e5fe5a868d738fc2a90372b8e288

Observation f70c6676-9a38-4cb1-9d3d-3cce8a6d8c35 · outbound

This paper cites Medical Image Analysis 58, 101562 (2019).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Medical Image Analysis 58, 101562 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.205494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.205494Z digest=sha256:5e7ac6e11ba535cb2d2de64a6e3346cf1f4f39952b1414599243df4c7aba4148

Observation 78361df4-daa6-42ef-b0aa-f166aacf199c · outbound

This paper cites On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.410077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.305953Z digest=sha256:be52bc7e990f782f473d1a0d5c109aaec1a4b30bce2fb53e1ce7aa55fdc93eea

Observation 2ee9eb48-4ba0-4738-8125-fca4412923ce · outbound

This paper cites In: ICML (2021) 4, 9, 19.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML (2021) 4, 9, 19

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.715955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.429736Z digest=sha256:629641bbff5bf51b0e1a22a6f75cfc0d0d5993cdcbf27cc639f34129755653f0

Observation 245362ca-8fb4-49ea-a018-1636c0a2a0eb · outbound

This paper cites Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.529248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.528599Z digest=sha256:61fa50beb6e95efbbc6a726b872f6a4b4d8e27fdde7ba77e9c24ad2f502a89d5

Observation 0e4732d4-e1a2-4b31-a9eb-378084a98ea3 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.339789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.597170Z digest=sha256:b76cefd0412d353a5ac8eda451ff74478532f4990c069ea8245721b316e04efe

Observation 6fdd2f16-5734-48ce-a493-ac6b85ea37e4 · outbound

This paper cites MedCLIP: Contrastive Learning from Unpaired Medical Images and Text.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedCLIP: Contrastive Learning from Unpaired Medical Images and Text

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.695492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.695492Z digest=sha256:0b0f562c6fe8bd259cc20802aac7a4f6fa90d9edd383850ded263345e32ecf83

Observation e30919d9-7ff2-4ffb-9a2e-df4ebc6656f8 · outbound

This paper cites an unresolved cited work.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.803512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.803512Z digest=sha256:da993d06dc0e3eac6f89803a9df06df905ff1122339fc23aa58ec6ff40f9cb81

Observation 8f4f9e26-01f8-41f2-83ee-40643a67c91c · outbound

This paper cites Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.173875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:20.895536Z digest=sha256:50f064816b1ba76f63384e3165c569fe4735dcf22f94819ba5a163b0d9c22c0a

Observation 69be5949-9051-4277-b855-3fa3e2acfa69 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.976948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.976948Z digest=sha256:e4a993ef93671361ffba06da5c806ddd7230b8146d09cc12da6d61ec08f951b7

Observation f2ecfa22-c5ac-4eb4-9fb4-792d2754a503 · outbound

This paper cites In-Distribution.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In-Distribution

Reference 30

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T15:21:22.036992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:21:21.068140Z digest=sha256:e748100cd1124d9e1e59a740cd24228faf1b3f8302a7f56b57bbf8c303755baa

Pith citing papers

Observation e131d344-9f9e-46f7-9fe1-dff80d68d595 · inbound

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift cites this paper.

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:13:53.947203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:13:53.947203Z digest=sha256:3ae6c39220d5e11a7ec9729f184495a4c4f43734fe13089e162c6155994ebe02