Pith. sign in

Paper Citation Record · LEDGER

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

As of 18 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2505.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15425 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:21.068140Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:13:53.947203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f9830d4a-77cd-49dc-bb44-3a02a7bea89d · outbound

This paper cites https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.441969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:17.569003Z digest=sha256:c6dcf1c4c31bc150e1905b7e32050f31975e8a5fcda5d21b9451d58cb723e8b5

Observation ba554940-a11b-4daa-adf3-e009f5bc99d8 · outbound

This paper cites A Survey of Medical Vision-and-Language Applications and Their Techniques.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? A Survey of Medical Vision-and-Language Applications and Their Techniques

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.662259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.662259Z digest=sha256:b5e2487ead6c984b0c44b5badfdfd9cdd0c852f006345a7e1c99463a7a690982

Observation bf83a1f4-5215-408a-b498-db43f58fb80e · outbound

This paper cites In: ICML Workshop on Computational Biology (2021) 2, 9, 20.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML Workshop on Computational Biology (2021) 2, 9, 20

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.260674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:17.808789Z digest=sha256:d921f4bedd2895417cd69aed3e21b52b87c6ea96835fc28fa36a4922afcf59d2

Observation 915abbaa-6cd8-407d-94a5-815b401de255 · outbound

This paper cites Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.850405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:17.873767Z digest=sha256:22434bd4d65a247c1c7b4b747569f41404ecad1df4e62d271f0d62533532c23e

Observation 33df9c7a-f727-447c-b8b0-3f4c76bdfce1 · outbound

This paper cites MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.967185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.967185Z digest=sha256:175731f784d76f0f84c90eb8010c0cfb57da7a36e41a530d24502ca7a578ae04

Observation 43739308-acd1-498c-aec4-5ee90fcd9201 · outbound

This paper cites In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.046793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.102723Z digest=sha256:37fa315754201090cfbe8c8c1d1645cf1643d2a24c629ad4d17fda17c1d2ae96

Observation 90a6e254-9ef8-4767-a418-54fd3f10fecd · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.860908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.204355Z digest=sha256:c77982cb75598f4605697fb8ea64d9fcc2c2f943b12181882b3c81cd28e16a94

Observation 956ac906-776d-4e26-aaf0-f4fd2f77d46c · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.653457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.321901Z digest=sha256:e3d73be023a2de08d3c9dc3c50fe30298b792507927b419f688996d23b05d6cf

Observation c53fdf87-3b86-48f8-91ef-d0c64306771a · outbound

This paper cites MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.675948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.499866Z digest=sha256:d3e86ee4bd2b46aa58f2eff484fe753c4271d81b920820b95080557c02397dc6

Observation 58959005-40ed-403b-83fa-087010bf6644 · outbound

This paper cites In: ICLR (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICLR (2019) 2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.434967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.635490Z digest=sha256:78fa3476487851624724938c17c122ee2bc66f7e0a41c90ff0d4ede586808e3b

Observation c6332889-c3c1-4d21-a701-e48ba1998e42 · outbound

This paper cites ICLR1(2), 3 (2022) 8.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? ICLR1(2), 3 (2022) 8

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.253117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:18.776022Z digest=sha256:c8c8a107ebcc4a5e3eaf9eeb9e5a510cdd171b2ba31adccbbcf2e794168b2510

Observation 4ef916d3-dec1-4414-b183-18395aa5422f · outbound

This paper cites Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:18.969671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:18.969671Z digest=sha256:13b3d343e0b7f174a55e73c7a58aa8f25a76394d0d2d0da70d367f020db52fff

Observation bd4549b0-3e2a-421f-bf14-7c5a85841c67 · outbound

This paper cites Noise is an Efficient Learner for Zero-Shot Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Noise is an Efficient Learner for Zero-Shot Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.132444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.132444Z digest=sha256:fce30196b1e773c7bdb1452b00ce477ab5512c10e927cce29331c5259fbdf911

Observation f675dad8-f92c-4e60-b223-63bef8bb45ef · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.961842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:19.292556Z digest=sha256:f28029518dd5342280b8b8f626943136ba6004579dfcf1dd9bf0c65b894fd2e8

Observation dc13bbbe-e59a-4483-92ac-88d0c26f5846 · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.434663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.434663Z digest=sha256:9d70d89f5c236fbe5d847b9cde9373dabbfb63a5059d57715997288a9a777cd1

Observation 1ef37400-0521-4c12-a2e3-0e6c5992fbd1 · outbound

This paper cites Radiology p.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Radiology p

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.735407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:19.547418Z digest=sha256:d3f12cade035d7dfc7b8003f31419c60153622e99bf7643ba55f3b2d96e050e4

Observation add91f25-6bb1-405f-bc53-8dc00554c113 · outbound

This paper cites IEEE Reviews in Biomedical Engineering (2025) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? IEEE Reviews in Biomedical Engineering (2025) 2

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.484419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:19.662098Z digest=sha256:8c1cc067e910c13cf4ebac24eb51977d281bed29720a1d9ba4616cf5b5f94114

Observation 85331b14-0378-4f8b-b6ba-70500ef184b5 · outbound

This paper cites UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.892988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.892988Z digest=sha256:5ea2a242f0f415863734e6c0241418d2304e5511e10232cec6f15e9daf9ca9d3

Observation fcfb872b-e090-4485-911c-a27521c16616 · outbound

This paper cites Journal of Biomedical Informatics135, 104234 (2022) 4.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Journal of Biomedical Informatics135, 104234 (2022) 4

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:19.989329Z digest=sha256:97a0cd583b8e81acf1ed743e7433c55c3867de4b5b1d617f945e815daf9d4c95

Observation d36f2d04-e14a-40b2-96e8-927b071b9a66 · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.936325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.112534Z digest=sha256:55d51bccc3e0d40f68c4b8185a662fac1469aeb1da25a2287b6eb1394e34ade2

Observation f70c6676-9a38-4cb1-9d3d-3cce8a6d8c35 · outbound

This paper cites Medical Image Analysis 58, 101562 (2019).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Medical Image Analysis 58, 101562 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.205494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.205494Z digest=sha256:15290ebb505410034cc612d7e3b691a2517b406e229eea83bb7e317076cf7910

Observation 78361df4-daa6-42ef-b0aa-f166aacf199c · outbound

This paper cites On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.410077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.305953Z digest=sha256:b3eb46cd4bf8240a186d81de7bede8cbbbb5b4bd6969f420e04baa168968c0a0

Observation 2ee9eb48-4ba0-4738-8125-fca4412923ce · outbound

This paper cites In: ICML (2021) 4, 9, 19.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML (2021) 4, 9, 19

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.715955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.429736Z digest=sha256:4c6ee4d97f78fee9e5a1e307ec0a537e63fb264b76554ec108b2368267996717

Observation 245362ca-8fb4-49ea-a018-1636c0a2a0eb · outbound

This paper cites Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.529248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.528599Z digest=sha256:4ad72c13bfb7652822c52a9ccd805d6d63541eccf0139ad5c093320d345edde6

Observation 0e4732d4-e1a2-4b31-a9eb-378084a98ea3 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.339789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.597170Z digest=sha256:b601fd93c6e46b66a26ad075fda649904d487c48e487c09113f1013932c62e3d

Observation 6fdd2f16-5734-48ce-a493-ac6b85ea37e4 · outbound

This paper cites MedCLIP: Contrastive Learning from Unpaired Medical Images and Text.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedCLIP: Contrastive Learning from Unpaired Medical Images and Text

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.695492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.695492Z digest=sha256:c941457de8a34fd19d73e1d89bd8a44589bf4ca0173e6f57fcb86e15c22ab41e

Observation e30919d9-7ff2-4ffb-9a2e-df4ebc6656f8 · outbound

This paper cites an unresolved cited work.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.803512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.803512Z digest=sha256:77cbac4eb17d188ab3d3325a1f561ce355e2479a375cfea43240369d1618ce5f

Observation 8f4f9e26-01f8-41f2-83ee-40643a67c91c · outbound

This paper cites Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.173875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:20.895536Z digest=sha256:49afe870d17168cc4f8eb27a1c0cf74539c7bf879aa64d3a6ddb5320557bcb93

Observation 69be5949-9051-4277-b855-3fa3e2acfa69 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.976948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.976948Z digest=sha256:9df7adeed212013e73a7e81918f5c2f200176b021aecd469303df3c549fc8ca6

Observation f2ecfa22-c5ac-4eb4-9fb4-792d2754a503 · outbound

This paper cites In-Distribution.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In-Distribution

Reference 30

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T15:21:22.036992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:21:21.068140Z digest=sha256:748a7ac91e0f799f8d243fc8df98d58879fc73daea504cdfb28741c7dfeb4079

Pith citing papers

Observation e131d344-9f9e-46f7-9fe1-dff80d68d595 · inbound

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift cites this paper.

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:13:53.947203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:13:53.947203Z digest=sha256:a88d0946b7b1258c54a6a09780a177cde0979dcb3ee29c6c03ad7b2ff5c72d24