Pith. sign in

Paper Citation Record · LEDGER

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

As of 9 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2505.15425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15425 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:21:21.068140Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:13:53.947203Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f9830d4a-77cd-49dc-bb44-3a02a7bea89d · outbound

This paper cites https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? https://pmc.ncbi.nlm.nih.gov/, accessed: 2025-04-05 9

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.441969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:17.569003Z digest=sha256:e98608c5574dc2e79f568e53930ffdd3c327f6dcff835d9467c268212a1a9271

Observation ba554940-a11b-4daa-adf3-e009f5bc99d8 · outbound

This paper cites A Survey of Medical Vision-and-Language Applications and Their Techniques.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? A Survey of Medical Vision-and-Language Applications and Their Techniques

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.662259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.662259Z digest=sha256:3d1bcbcf5040c8c5d1a958858d866d9f8286866958ca2bdec05efdd20246430e

Observation bf83a1f4-5215-408a-b498-db43f58fb80e · outbound

This paper cites In: ICML Workshop on Computational Biology (2021) 2, 9, 20.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML Workshop on Computational Biology (2021) 2, 9, 20

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.260674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:17.808789Z digest=sha256:768bf3792cb8067dac23f0c8922d0c77bf64c907e692bc482c4c834a51276da3

Observation 915abbaa-6cd8-407d-94a5-815b401de255 · outbound

This paper cites Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.850405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:17.873767Z digest=sha256:622171d3b8269b8522b0224939fc8b9cd575af89f26aa3809238722477b1f4a7

Observation 33df9c7a-f727-447c-b8b0-3f4c76bdfce1 · outbound

This paper cites MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedMNIST-C: Comprehensive benchmark and improved classifier robustness by simulating realistic image corruptions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:17.967185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:17.967185Z digest=sha256:da7d8373b337fb2c055ce95823faa9faf99fadcc9edd879278c4c3b84a429633

Observation 43739308-acd1-498c-aec4-5ee90fcd9201 · outbound

This paper cites In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: 2025 IEEE 22nd In- ternational Symposium on Biomedical Imaging (ISBI)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:25.046793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.102723Z digest=sha256:dba7955d932ab72d30538a7050f7864c7aa9ac5aea79f6d00204b1bba880bdf3

Observation 90a6e254-9ef8-4767-a418-54fd3f10fecd · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.860908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.204355Z digest=sha256:98f76885d9e3872bc4d17f221a9f5d716625745f46b12ddb8e813fae282f56b9

Observation 956ac906-776d-4e26-aaf0-f4fd2f77d46c · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.653457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.321901Z digest=sha256:884443f9689628f62d270ad31710f5027a7adce823430ddd23c7b742aa5d59e0

Observation c53fdf87-3b86-48f8-91ef-d0c64306771a · outbound

This paper cites MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedFuse: Multi-modal fusion with clinical time-series data and chest X-ray images

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.675948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.499866Z digest=sha256:106a35826111c87d250d834c70c6c41243040f923f38cc0791b2c0d3b3617d34

Observation 58959005-40ed-403b-83fa-087010bf6644 · outbound

This paper cites In: ICLR (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICLR (2019) 2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.434967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.635490Z digest=sha256:fe786f0d91b3aa7773e963d895d1138fe803c53fb04db0fe792851e526347b33

Observation c6332889-c3c1-4d21-a701-e48ba1998e42 · outbound

This paper cites ICLR1(2), 3 (2022) 8.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? ICLR1(2), 3 (2022) 8

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:24.253117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:18.776022Z digest=sha256:69f1a751ae661a90e6e631a7abdb6c226ec5fe736c1a9aee0a3a264b5d53bc08

Observation 4ef916d3-dec1-4414-b183-18395aa5422f · outbound

This paper cites Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:18.969671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:18.969671Z digest=sha256:78071b9ed4a92ee7cdeb6ab20e2b746e3a4812552b25d272274775536534e8c0

Observation bd4549b0-3e2a-421f-bf14-7c5a85841c67 · outbound

This paper cites Noise is an Efficient Learner for Zero-Shot Vision-Language Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Noise is an Efficient Learner for Zero-Shot Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.132444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.132444Z digest=sha256:94e2941a5976f977b51368c7aa4b2687a0cc476d10756658fcbc3fa214cddf30

Observation f675dad8-f92c-4e60-b223-63bef8bb45ef · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Proceedings of the AAAI Conference on Artificial Intelligence (2019) 2

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.961842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:19.292556Z digest=sha256:eb08d8b4bf1398db793b54597e0c4b05efabd151887c7ec55785e09be497982b

Observation dc13bbbe-e59a-4483-92ac-88d0c26f5846 · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.434663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.434663Z digest=sha256:4e6108ed891213622e8d5e6144ae6aeef8fe5bccc0b780bc4c664d4203e697f0

Observation 1ef37400-0521-4c12-a2e3-0e6c5992fbd1 · outbound

This paper cites Radiology p.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Radiology p

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.735407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:19.547418Z digest=sha256:d151eb022c8821c8ab71b85b391cc91f8026a4f14931adc6ebd39d06ced3e92e

Observation add91f25-6bb1-405f-bc53-8dc00554c113 · outbound

This paper cites IEEE Reviews in Biomedical Engineering (2025) 2.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? IEEE Reviews in Biomedical Engineering (2025) 2

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.484419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:19.662098Z digest=sha256:5ec2524bba0d0a11d5d6a8196fc11c90ef01e8ec825db7c29e32a0aa36ce6940

Observation 85331b14-0378-4f8b-b6ba-70500ef184b5 · outbound

This paper cites UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? UniMed-CLIP: Towards a Unified Image-Text Pretraining Paradigm for Diverse Medical Imaging Modalities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:19.892988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:19.892988Z digest=sha256:169fcc1472d007f8c13aa2538a93e195b12280e6a68ed9b9398532e55947eede

Observation fcfb872b-e090-4485-911c-a27521c16616 · outbound

This paper cites Journal of Biomedical Informatics135, 104234 (2022) 4.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Journal of Biomedical Informatics135, 104234 (2022) 4

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:23.210950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:19.989329Z digest=sha256:20861160818ddf57ae757d0a0b4c3d2179e9267dccdc72967ea81b66db8db206

Observation d36f2d04-e14a-40b2-96e8-927b071b9a66 · outbound

This paper cites In: International Conference on Medical Image Computing and Computer-Assisted Intervention.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: International Conference on Medical Image Computing and Computer-Assisted Intervention

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.936325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.112534Z digest=sha256:597ee1a11f9555690ac4c587a0d5c928212d8794fa2a518cee2592b13bb1eaf0

Observation f70c6676-9a38-4cb1-9d3d-3cce8a6d8c35 · outbound

This paper cites Medical Image Analysis 58, 101562 (2019).

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Medical Image Analysis 58, 101562 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.205494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.205494Z digest=sha256:5e7ac6e11ba535cb2d2de64a6e3346cf1f4f39952b1414599243df4c7aba4148

Observation 78361df4-daa6-42ef-b0aa-f166aacf199c · outbound

This paper cites On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Models

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:21:21.410077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.305953Z digest=sha256:b271de4a45cba837c8e99d03ee868990682af4286239ff19890d58898f04c520

Observation 2ee9eb48-4ba0-4738-8125-fca4412923ce · outbound

This paper cites In: ICML (2021) 4, 9, 19.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: ICML (2021) 4, 9, 19

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.715955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.429736Z digest=sha256:a5ece51cd25b85b875896d50fce19659badaacb4dc340445be21305cc7cde6b0

Observation 245362ca-8fb4-49ea-a018-1636c0a2a0eb · outbound

This paper cites Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Seco de Herrera, A., et al.: Rocov2: Radiol- ogy objects in context version 2, an updated multimodal image dataset

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.529248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.528599Z digest=sha256:2a1728ce5dcc823e36d97d5f78b4fe91cc2ab3645d6073ac7c205995c8c4df5e

Observation 0e4732d4-e1a2-4b31-a9eb-378084a98ea3 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.339789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.597170Z digest=sha256:a9169e7ce1ec464d2218e88ebfed2506d4bf254107c16686b026946c583a712a

Observation 6fdd2f16-5734-48ce-a493-ac6b85ea37e4 · outbound

This paper cites MedCLIP: Contrastive Learning from Unpaired Medical Images and Text.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? MedCLIP: Contrastive Learning from Unpaired Medical Images and Text

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.695492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.695492Z digest=sha256:0b0f562c6fe8bd259cc20802aac7a4f6fa90d9edd383850ded263345e32ecf83

Observation e30919d9-7ff2-4ffb-9a2e-df4ebc6656f8 · outbound

This paper cites an unresolved cited work.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.803512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.803512Z digest=sha256:da993d06dc0e3eac6f89803a9df06df905ff1122339fc23aa58ec6ff40f9cb81

Observation 8f4f9e26-01f8-41f2-83ee-40643a67c91c · outbound

This paper cites Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? Frontiers in neuroinformatics13, 46 (2019) 2, 4, 5

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:21:22.173875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:20.895536Z digest=sha256:b12bdedfeacd5a5a3c2eeae4d9d874e3c24361d374b4760cfa467c9ecedfb049

Observation 69be5949-9051-4277-b855-3fa3e2acfa69 · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:21:20.976948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:21:20.976948Z digest=sha256:e4a993ef93671361ffba06da5c806ddd7230b8146d09cc12da6d61ec08f951b7

Observation f2ecfa22-c5ac-4eb4-9fb4-792d2754a503 · outbound

This paper cites In-Distribution.

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable? In-Distribution

Reference 30

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T15:21:22.036992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:21:21.068140Z digest=sha256:6d50355d996e1d7b42a9b7873d1c88cdfe25ecf77deaf68ae4531cdae126e617

Pith citing papers

Observation e131d344-9f9e-46f7-9fe1-dff80d68d595 · inbound

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift cites this paper.

Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T19:13:53.947203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:13:53.947203Z digest=sha256:3ae6c39220d5e11a7ec9729f184495a4c4f43734fe13089e162c6155994ebe02