Pith. sign in

Paper Citation Record · LEDGER

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models

As of 8 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2506.08990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08990 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:21.165503Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9fd25426-9bcc-4eb4-bee7-b13f40e3ed74 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Imagenet: A large-scale hierarchical image database,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.773766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.773766Z digest=sha256:e3e3a7036ad159c58f34104b2d71e9d5f35d9fde004d31e471f7e8726c4c282a

Observation 247791b4-f9dd-4f94-8849-25c385551f2c · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.823905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.823905Z digest=sha256:7b31d102a98fdc40ad429960ce88d681fc0730e44d985a8848fec905a9e28507

Observation 9de23ca4-edcd-4f77-a29c-69e2200a2b5a · outbound

This paper cites Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:22.003020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.873856Z digest=sha256:d086254571cd4303088c9bec5886b21883aca58fbe057237140742a582f7f0cc

Observation dad709e1-77fc-482f-9a7b-e01b3d739726 · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.986416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.929274Z digest=sha256:954ae47a5020fb3ada12f69c3e4edba91152d5f5586ddb9056eec5d8dd3e5f0d

Observation e8d136c1-2ca1-422c-9dac-c5d1403ff169 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning transferable visual models from natural language supervision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.949952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.949952Z digest=sha256:8dee07d430604361df769fc5c0e824a0b7b387e8b6dc22a375ec7843759590cd

Observation 76f20c04-0ca6-4e17-98a2-cea9eab6695c · outbound

This paper cites Contrastive learning of medical visual representations from paired images and text,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive learning of medical visual representations from paired images and text,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.961508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.954272Z digest=sha256:a513deff984339efee19a80d0932a858d977036131cf541e0a774d20f5b76385

Observation 72dc1b6d-e7a1-497c-824a-7adabb02216a · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.948018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.959477Z digest=sha256:bb26b6a047bac55d2ec4a89edf769597f6455d0b9f7576d6a45c93ec3ed40ba3

Observation 45fb1f3a-ccf5-48ad-995d-7bbbca04886e · outbound

This paper cites Making the most of text semantics to improve biomedical vision– language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Making the most of text semantics to improve biomedical vision– language processing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.963272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.963272Z digest=sha256:19729e23d00d51c985846fb728c61d802e5803714c697605151b28b0bfd30688

Observation e3ef6427-e5b4-4135-9c46-d39e5e22dc4f · outbound

This paper cites A multimodal biomedical foundation model trained from fifteen million image–text pairs,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models A multimodal biomedical foundation model trained from fifteen million image–text pairs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.923279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.967837Z digest=sha256:d3d76545b895bbbbc7a05a278f9e6f7c373cf10db9408a194448fc06e17af792

Observation ac4bc219-aaac-4734-8b41-47db5935ee65 · outbound

This paper cites Advancing radiograph representation learning with masked record modeling,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Advancing radiograph representation learning with masked record modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.909308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.971917Z digest=sha256:e27c411ddb2f6666a2b0195d366ae5d6a2d554b0c9d4b98131a66c9cab0cbfb3

Observation 1944ea8f-beef-46cf-a372-fdb3f8bab1c2 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.975873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.975873Z digest=sha256:1b773559f4f7b85d7c67077ff8b23a15c2d226e0730b4987e755fe05df931bb5

Observation cf92113c-2d18-48dd-ab23-d108399a51d8 · outbound

This paper cites Medical Vision Language Pretraining: A survey.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medical Vision Language Pretraining: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.981003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.981003Z digest=sha256:4c08880b1f3a5d327524e7af0fbd5d59ab4b382ff4e9c94c66cdf77a8743a88c

Observation 1e88a606-8554-46d1-b82d-162313e4065c · outbound

This paper cites Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.895563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.985555Z digest=sha256:5519a3b0794e8a1f96b8232d61e2428877570583d34878f041599b40ca8ef2b8

Observation 36067afd-f3b0-47bd-aa4f-44169d599346 · outbound

This paper cites Im- proving medical vision-language contrastive pretraining with semantics- aware triage,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Im- proving medical vision-language contrastive pretraining with semantics- aware triage,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.881471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.989976Z digest=sha256:531f1d9be300471b95ada9ce5f3c1ae998f28fd576676ae80912f471911e244d

Observation 6d72bb59-ce72-44a4-9a8c-fad5114fa252 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.867278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.993776Z digest=sha256:e674cae34c7a6340562ecd3bb36ffc6933cb89b8c2d0785f2d6eacaeac6eb7cd

Observation 9d5e6252-eb2a-4858-b259-dda8df479418 · outbound

This paper cites Learning to exploit temporal structure for biomedical vision-language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning to exploit temporal structure for biomedical vision-language processing,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.998341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.998341Z digest=sha256:fc0aa51a7135aeaae8c6d64f0f13059e78b018f5ef06aa95e82fca52275ad22a

Observation 984b26ad-05cb-41e8-b24a-31d5bad53fe8 · outbound

This paper cites Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.842083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.002149Z digest=sha256:96da54341c228e82905a88c399e3dca537d0efd8f8842eaf23c310305d0ca4b1

Observation 331a124b-a0f9-40ab-be09-e1410669366b · outbound

This paper cites Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.826179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.006246Z digest=sha256:bc29e8ab47fc7f6f9787403ea59fa6d66ad4cee03324e36b43f73d44d11729eb

Observation 2390c2ce-d423-4c67-9e85-c087b41cd959 · outbound

This paper cites Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.809530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.010139Z digest=sha256:eb2e3e8efa387fcd5e5dc1d29b9b8b4e2c4f5dce79c33e86e191e6a37f8624aa

Observation 9a6e642a-3f72-4cf6-89db-62c0af89edad · outbound

This paper cites Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.793117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.014390Z digest=sha256:f65793ae74e42890baa6821bc54e2fad81620a845236740dfc32d05b119fd2a4

Observation 97f3da43-ea3f-477b-a633-d5d0676e25b3 · outbound

This paper cites Deltanet: Conditional medical report generation for covid- 19 diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Deltanet: Conditional medical report generation for covid- 19 diagnosis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.775166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.019327Z digest=sha256:aab19259bcd525da6010477952316a36411fa92967963f3738aa2d4f87c812bd

Observation 67f3d0d5-0c5d-4ed6-ab47-b094e5d444cc · outbound

This paper cites Self-supervised learning for medical image analysis using image context restoration,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Self-supervised learning for medical image analysis using image context restoration,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.023679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.023679Z digest=sha256:316afe89bc8ae1873d2753e24731a7d1ecde7fc2f51e81c0bc06ae339f98c0cd

Observation 9266745b-4665-4de7-9dab-56ec3fdc579f · outbound

This paper cites Models genesis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Models genesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.745864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.027902Z digest=sha256:a033c23de069696b33bdbcd6911e245ca30a05d58225e0510262e4b9ef86437e

Observation 64b13a44-3f32-4af8-b3c9-a5ec1d543910 · outbound

This paper cites Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.728548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.032152Z digest=sha256:1bb9976f62bd3481ea97e7aea5510d4bfd613ecfb4116832332e8a938fc48755

Observation 092cee43-8f44-407b-85d3-08a0d796bb1e · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.035993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.035993Z digest=sha256:11f5eb396437c3cb8b7dd48addcfc450f4b0da53470388d90f24a827daf50253

Observation 825e5557-90db-47c7-8405-3d29a0b86a88 · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.700519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.040046Z digest=sha256:cbd0994cff9956ffe07156431c3912788b1a6cab4ce9b8f333d8783842b2cc93

Observation da009c55-246b-49ff-88d3-e3ca112e52dc · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Align before fuse: Vision and language representation learning with momentum distillation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.044391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.044391Z digest=sha256:387e3ffa88170fe6df3664b6bcb85cdc88cdef55332f2761a0b02a3f9eb07590

Observation 31c42f06-88bd-4af9-8adb-30ca8ec081aa · outbound

This paper cites Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.048709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.048709Z digest=sha256:58e08b3f518a369677506ca3687f40759cd8c919b4aa5371eca848c4df014a65

Observation 6c061e64-054d-4aa8-a068-8b045d806f30 · outbound

This paper cites Knowledge- enhanced visual-language pre-training on chest radiology images,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Knowledge- enhanced visual-language pre-training on chest radiology images,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.672655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.052954Z digest=sha256:ef134bb7b16246447a1d7d97ae7b2f0565f605f430576a3bbe17ce01ab08c884

Observation 82213d81-baba-48b2-bfa3-c8891ca4ba0b · outbound

This paper cites Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.056753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.056753Z digest=sha256:48f7538a7c5f117cded9e3f0d7898a9482eca0092aca024665f397b2bc2a3e8a

Observation f7a18f76-d20c-4853-af0a-549b7b906163 · outbound

This paper cites Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.643981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.060464Z digest=sha256:8be5180a861e56343ee4f1b8b773042dfbaeaedbd9240866d692aed4cf20d301

Observation cf44f368-257a-4fed-8d31-577fdf9d7fc8 · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.064175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.064175Z digest=sha256:872c46b72031c463ec5df1c1631fdb5800eb820c90a8cc528b3ac3739ab230f4

Observation e597b183-f5b7-4d40-becd-542f72198e17 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-efficient transfer learning for nlp,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.068498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.068498Z digest=sha256:09eb4260edea0bf0ee7c55532d074653ac1b06948f370bcff77980e60ae95a5f

Observation 4dd1dd1f-afbf-4c0f-9aed-886cfd4e1571 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.072896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.072896Z digest=sha256:e791e1e847fe0107f4abc7b10ff6bb1dc7e7c60632febc3423b1ca34c088f793

Observation 2414083d-5b26-41e4-b267-929771e16a86 · outbound

This paper cites Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.600504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.077066Z digest=sha256:cd95291de155e09ddb07e1d711a83b2a426ed5fbf361755d5eeb7c85309f255c

Observation adf02300-4504-4571-be33-5c3b7c460cf1 · outbound

This paper cites Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.582459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.081187Z digest=sha256:cb0f7555f8382186625e18988178dea4ffcf0e27f0be18c6ae44845c354f7cc9

Observation 582f3bc9-43a9-4b30-80ad-3c3bb71c53fa · outbound

This paper cites Towards a Unified View of Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards a Unified View of Parameter-Efficient Transfer Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.085120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.085120Z digest=sha256:3b5aaddc7bc70bb66b647c4fa47f0719a41d34c5989c6ceb4db8a9925c59f26f

Observation 6d08116e-f5cd-4dc2-9991-c6c6cbcf7a3d · outbound

This paper cites Visual prompt tuning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Visual prompt tuning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.089229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.089229Z digest=sha256:26dd6a134591b96077cd055acdbeed5c631f99a4680f092577aab0245db6f581

Observation 9cb90014-3299-4f7d-bbe1-e5906504b314 · outbound

This paper cites Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.557861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.093214Z digest=sha256:3c5142d90c933bf4ce8a6d90f8e84fb44845fa84eddb018c980070c00509f512

Observation cd240235-3b9b-41ee-b783-60e85f662059 · outbound

This paper cites AIM: Adapting Image Models for Efficient Video Action Recognition.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models AIM: Adapting Image Models for Efficient Video Action Recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.096816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.096816Z digest=sha256:b96d1f1b87e2613ea7885357ce4899e39951fd3e90ee2d6799b7c9b45e1b6cbe

Observation b1b5eee8-d624-4a21-8933-82b1f20b639e · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vision Transformer Adapter for Dense Predictions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.100923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.100923Z digest=sha256:2f5e22d88ae9aff17f372645b92d4a6f7cd031ce377879de208d7989def391dc

Observation f0189b50-bdb5-4c9c-b43f-64fff44af3f5 · outbound

This paper cites Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.105209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.105209Z digest=sha256:c7d3fb1af29b302a4d7d8928b32f74f26987b88a2df17e0d7d9591a2a0f8a546

Observation 3cf57145-50b6-4376-b42d-feb2aa78d11f · outbound

This paper cites MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.109848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.109848Z digest=sha256:fbe58e179576ca8e0753a6e74ecb25ba6062ec4afcdb0079df808e3331c32cfa

Observation 940aa1de-6d6a-46ff-a5c0-aaa6968df2bc · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.115036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.115036Z digest=sha256:ee90388504b1f868ed9b37b0f494e3ab9932a3a34275c830bf50782f3584a2fc

Observation 9c12a0b7-ad35-4518-b3c9-b359c796cb09 · outbound

This paper cites Prompt tuning for parameter- efficient medical image segmentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Prompt tuning for parameter- efficient medical image segmentation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.542188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.119838Z digest=sha256:a764f16121d569b0271cc45df8b871a98f40f0c5f0adc0340a39c89b5eaa8b22

Observation 12073882-8e25-47fa-b8f3-13c49385940a · outbound

This paper cites Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.525535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.123971Z digest=sha256:16ae1b9d4ef06cf322a91555314fc987a804bfc78071164bd336b1a4c8f8be1e

Observation fd6089af-958e-43a2-98df-a43146bd29da · outbound

This paper cites Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.128128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.128128Z digest=sha256:9c2e6f3a1d649476aa8ba954d1ef306472f6063855558f48b60bade65dad5196

Observation dfda98a6-9006-4d5c-a902-26e0b1555765 · outbound

This paper cites Scaling language- image pre-training via masking,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Scaling language- image pre-training via masking,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.509160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.132633Z digest=sha256:1df242f5d3039f47fae0599b1cbc6d978a77c123fa648eed4661dd207a09ca38

Observation 677245e6-b7da-4913-b2a4-b271c50526f2 · outbound

This paper cites Attention is all you need,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Attention is all you need,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.136469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.136469Z digest=sha256:6b9deb296ab7efe50c6f1bfa5c1f4842c1f2c9f9592828ed576086e7698a57e6

Observation e36c3afa-c4bb-4bf9-af8b-6324a2c96d89 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gaussian Error Linear Units (GELUs)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.140339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.140339Z digest=sha256:acf76fa9bcc241d5ce1213d44980a62c0218275832a40cc168f974164a737942

Observation 56ec48d9-6ab1-429d-8320-7828089505f6 · outbound

This paper cites Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.144569Z digest=sha256:2a66feef7e6a2a41c78880bc50ce78ffb480d12a28d1ee89b019cffde4d6884b

Observation 88e534ca-45b7-4aef-b061-c2390706973e · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Representation Learning with Contrastive Predictive Coding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.149064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.149064Z digest=sha256:7dc60262f07ceff08909d509ca4e2918367d4a3fdeeb40e34b90ba0f6584e3b7

Observation ed74d849-4d17-493c-90d0-3a5ca81c09aa · outbound

This paper cites Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.470237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.153308Z digest=sha256:2b4378959b54a6e8565f6eaace96a33de29ec71c835d3a0bb3a3cd73f90973fb

Observation c3d7a77f-16a5-4936-be96-df9a5d821719 · outbound

This paper cites Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.157303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.157303Z digest=sha256:a0f0ed3d64fc99179a5421a98b073e10beacc9c26f17b71c95539cb4fe7e1e57

Observation 01006229-228d-4ad5-9916-e3da28397fdb · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Pytorch: An imperative style, high-performance deep learning library,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.161272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.161272Z digest=sha256:59e7a36145e7b4c9eb18a7cd9c0f444d950247b1aefd216fd2ffe0a9fd6a9a29

Observation 75493d61-a82c-408a-8cd7-dfe7fe39055b · outbound

This paper cites Decoupled Weight Decay Regularization.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Decoupled Weight Decay Regularization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.165503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.165503Z digest=sha256:a4d51659445d1af6fb7163020293b74a48d421e4d56c389a93cef94b0e9bdb84

Pith citing papers

No inbound Pith citation observations are available.