Pith. sign in

Paper Citation Record · LEDGER

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models

As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2506.08990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08990 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:02:21.165503Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9fd25426-9bcc-4eb4-bee7-b13f40e3ed74 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Imagenet: A large-scale hierarchical image database,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.773766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.773766Z digest=sha256:072ed4ac52728113028a76a680218fa3efa11efd9df90c0a06db7660f1c86f97

Observation 247791b4-f9dd-4f94-8849-25c385551f2c · outbound

This paper cites MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.823905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.823905Z digest=sha256:b72cad8ae9afd6bbb51d23278ce7a8237b379c12aae4c4521b829aae19237f60

Observation 9de23ca4-edcd-4f77-a29c-69e2200a2b5a · outbound

This paper cites Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:22.003020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.873856Z digest=sha256:d1cc65d173fc94798f3b688d8d022e31250676508a5d9928d6e0eb12db358bd9

Observation dad709e1-77fc-482f-9a7b-e01b3d739726 · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.986416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.929274Z digest=sha256:3a00fa89cf4edc7e8f1a5a9b0dc2b923e43978065efc287931f2c68fb95977d6

Observation e8d136c1-2ca1-422c-9dac-c5d1403ff169 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning transferable visual models from natural language supervision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.949952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.949952Z digest=sha256:74162bfa44d807c232e670980e49074b1023ce63d778cd4c4c09e17bd24bf21b

Observation 76f20c04-0ca6-4e17-98a2-cea9eab6695c · outbound

This paper cites Contrastive learning of medical visual representations from paired images and text,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive learning of medical visual representations from paired images and text,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.961508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.954272Z digest=sha256:b97938259cc509c714c2a72398796be6a5083599ed7a2c9d73afdd9a0bb02ad2

Observation 72dc1b6d-e7a1-497c-824a-7adabb02216a · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.948018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.959477Z digest=sha256:d092e66eacee8ed308710a4b448566f6dddf2046e1df1fbb9477d1e2449fad27

Observation 45fb1f3a-ccf5-48ad-995d-7bbbca04886e · outbound

This paper cites Making the most of text semantics to improve biomedical vision– language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Making the most of text semantics to improve biomedical vision– language processing,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.963272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.963272Z digest=sha256:2b920aa9aea1030b12a4f5cb4097f71b88471369435df2ce23c43e4b3cef2f68

Observation e3ef6427-e5b4-4135-9c46-d39e5e22dc4f · outbound

This paper cites A multimodal biomedical foundation model trained from fifteen million image–text pairs,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models A multimodal biomedical foundation model trained from fifteen million image–text pairs,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.923279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.967837Z digest=sha256:7a733424e1e0d109ea622b69d38a016d60f324ad5b0b163f3be8813feefeb86e

Observation ac4bc219-aaac-4734-8b41-47db5935ee65 · outbound

This paper cites Advancing radiograph representation learning with masked record modeling,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Advancing radiograph representation learning with masked record modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.909308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.971917Z digest=sha256:18a5d3729cc4e83f2c52e44f2cde589fcefc10afe6977278921e1456c84e7fc6

Observation 1944ea8f-beef-46cf-a372-fdb3f8bab1c2 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.975873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.975873Z digest=sha256:99921733aa2319bbef0e61669897a5734bbf7640f6d81d24ba5686c1c1b00a45

Observation cf92113c-2d18-48dd-ab23-d108399a51d8 · outbound

This paper cites Medical Vision Language Pretraining: A survey.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medical Vision Language Pretraining: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.981003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.981003Z digest=sha256:a0df0fd368e0e89fd3a3effb8b552e537412ac8edf418575efcd23f75e6dba07

Observation 1e88a606-8554-46d1-b82d-162313e4065c · outbound

This paper cites Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Unichest: Conquer-and-divide pre-training for multi-source chest x-ray classifica- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.895563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.985555Z digest=sha256:12b33257ba03c2f031ae16573527e42f0715054211c703532df2c05af967af79

Observation 36067afd-f3b0-47bd-aa4f-44169d599346 · outbound

This paper cites Im- proving medical vision-language contrastive pretraining with semantics- aware triage,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Im- proving medical vision-language contrastive pretraining with semantics- aware triage,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.881471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.989976Z digest=sha256:cc7387ca8ee849a2b579ca0037c9747700fc3f06e27c5f74cc0b5ca586216a65

Observation 6d72bb59-ce72-44a4-9a8c-fad5114fa252 · outbound

This paper cites Exploring scalable medical image encoders beyond text supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Exploring scalable medical image encoders beyond text supervision,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.867278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:20.993776Z digest=sha256:fc2fe30cf0b02533ae49dd07b3e61402d6132cab47e257564328885b87526ac0

Observation 9d5e6252-eb2a-4858-b259-dda8df479418 · outbound

This paper cites Learning to exploit temporal structure for biomedical vision-language processing,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Learning to exploit temporal structure for biomedical vision-language processing,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:20.998341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:20.998341Z digest=sha256:d11f240d11e715501ad77820ad6b5d86fafb7cca4514fe61a3a66de001a776bc

Observation 984b26ad-05cb-41e8-b24a-31d5bad53fe8 · outbound

This paper cites Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gener- alized radiograph representation learning via cross-supervision between images and free-text radiology reports,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.842083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.002149Z digest=sha256:6f02c70c53f8030e3487ddc8735832b13d570c98897a050ee0d9d78b4ded5ccb

Observation 331a124b-a0f9-40ab-be09-e1410669366b · outbound

This paper cites Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexrelnet: An anatomy-aware model for tracking longi- tudinal relationships between chest x-rays,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.826179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.006246Z digest=sha256:fe714cae7d3621f3b46ad9a965f0ef3c8e0755ad42b9038d5fbbab9617ddec86

Observation 2390c2ce-d423-4c67-9e85-c087b41cd959 · outbound

This paper cites Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Chexfusion: Effective fusion of multi-view features using transformers for long-tailed chest x-ray classification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.809530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.010139Z digest=sha256:2f65c3b6a20418d6c983625b14065968a58d05f3e1deb4c1416e04fbd480de40

Observation 9a6e642a-3f72-4cf6-89db-62c0af89edad · outbound

This paper cites Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Act like a radiologist: towards reliable multi-view correspondence reasoning for mammogram mass detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.793117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.014390Z digest=sha256:05e9194be93b3f00d0a205140be57968f3632ff010c6733e1cc705ac4a562e3e

Observation 97f3da43-ea3f-477b-a633-d5d0676e25b3 · outbound

This paper cites Deltanet: Conditional medical report generation for covid- 19 diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Deltanet: Conditional medical report generation for covid- 19 diagnosis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.775166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.019327Z digest=sha256:b64d65058e63db7056f04a03781af928a22a5f3c67e9febe3ecc0c1151432085

Observation 67f3d0d5-0c5d-4ed6-ab47-b094e5d444cc · outbound

This paper cites Self-supervised learning for medical image analysis using image context restoration,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Self-supervised learning for medical image analysis using image context restoration,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.023679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.023679Z digest=sha256:28a66185910d2ea40424d3c8b3b9313969d894c55433e7536424a68c1b97da4f

Observation 9266745b-4665-4de7-9dab-56ec3fdc579f · outbound

This paper cites Models genesis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Models genesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.745864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.027902Z digest=sha256:3a63455ba871222ae41fb216f27b8376f8d0853cb4297d4538ad73dab0c7c749

Observation 64b13a44-3f32-4af8-b3c9-a5ec1d543910 · outbound

This paper cites Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Comparing to learn: Surpassing imagenet pretraining on radiographs by comparing image representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.728548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.032152Z digest=sha256:74725a5c74584ee7ab7e0daf10e3f19ed8983840083863edd24e6a2cea5d8cfe

Observation 092cee43-8f44-407b-85d3-08a0d796bb1e · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.035993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.035993Z digest=sha256:13d6f89dc63e15a375ea5be6fedab9f8e62b4cbbc6a827c50ad6a9b47e3721fc

Observation 825e5557-90db-47c7-8405-3d29a0b86a88 · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.700519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.040046Z digest=sha256:0a1f98688fe24e0ff5f298d76346ec136be3fda7c02820381126a44d5038fce7

Observation da009c55-246b-49ff-88d3-e3ca112e52dc · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Align before fuse: Vision and language representation learning with momentum distillation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.044391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.044391Z digest=sha256:69cb7db24303df3924676f5ebafd098fe2dbdbed6ce39f8ebca75ff151f7c6fe

Observation 31c42f06-88bd-4af9-8adb-30ca8ec081aa · outbound

This paper cites Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.048709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.048709Z digest=sha256:801ce7a089eb08b93a4405c9dbca65df1b705c2872a144d0067ad2924cad56ac

Observation 6c061e64-054d-4aa8-a068-8b045d806f30 · outbound

This paper cites Knowledge- enhanced visual-language pre-training on chest radiology images,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Knowledge- enhanced visual-language pre-training on chest radiology images,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.672655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.052954Z digest=sha256:24545810a23b6fe4ef391f5bf9380399d75b8bcb38e91bf96ce14ee5b9fae2e8

Observation 82213d81-baba-48b2-bfa3-c8891ca4ba0b · outbound

This paper cites Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Carzero: Cross-attention alignment for radiology zero-shot classifica- tion,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.056753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.056753Z digest=sha256:2ac6298ff4c9b6ae1451deee336c5c19746818f26b248652901398638b914b86

Observation f7a18f76-d20c-4853-af0a-549b7b906163 · outbound

This paper cites Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Enhancing representation in radiography- reports foundation model: A granular alignment algorithm using masked contrastive learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.643981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.060464Z digest=sha256:214df7594993298a82c233c9f71d1a5433ff177963a362084d27c9c05e1215d7

Observation cf44f368-257a-4fed-8d31-577fdf9d7fc8 · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.064175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.064175Z digest=sha256:9982cc9616462ec38f544f06bddd359f4351606672ce7b686a980fe3714668cf

Observation e597b183-f5b7-4d40-becd-542f72198e17 · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-efficient transfer learning for nlp,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.068498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.068498Z digest=sha256:6c935f5ae6b01186109d2064dc346172c589873d58d5f30095991b85a1ec94de

Observation 4dd1dd1f-afbf-4c0f-9aed-886cfd4e1571 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.072896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.072896Z digest=sha256:b0c27b814ee797c0428d748b160016e598a6686da0a89f2416ceacd0d4a2fb1d

Observation 2414083d-5b26-41e4-b267-929771e16a86 · outbound

This paper cites Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.600504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.077066Z digest=sha256:ec9c6c54ed0b599cee9604e48e127fc13ed06350ef55967ad78a6fd6f923a9d0

Observation adf02300-4504-4571-be33-5c3b7c460cf1 · outbound

This paper cites Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Bitfit: Simple parameter- efficient fine-tuning for transformer-based masked language-models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.582459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.081187Z digest=sha256:b69159702afce402fed674c6aa6ec1f4e99aa874fa456c1a3f8e3cc0841d138f

Observation 582f3bc9-43a9-4b30-80ad-3c3bb71c53fa · outbound

This paper cites Towards a Unified View of Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards a Unified View of Parameter-Efficient Transfer Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.085120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.085120Z digest=sha256:b5a5cd36df75b6216d1d4706db4f0647ca6003cd5317b3c1f01a020e87614eae

Observation 6d08116e-f5cd-4dc2-9991-c6c6cbcf7a3d · outbound

This paper cites Visual prompt tuning,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Visual prompt tuning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.089229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.089229Z digest=sha256:6a768b1d84d4d17cad3cd0f3de90d536818b21fcf7c18827d0eb0bc5c3ec985c

Observation 9cb90014-3299-4f7d-bbe1-e5906504b314 · outbound

This paper cites Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vl-adapter: Parameter-efficient transfer learning for vision-and-language tasks,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.557861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.093214Z digest=sha256:f570d376d147fa8187d1643b06c9d119f06271328f7cb348c1087f84bb8766df

Observation cd240235-3b9b-41ee-b783-60e85f662059 · outbound

This paper cites AIM: Adapting Image Models for Efficient Video Action Recognition.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models AIM: Adapting Image Models for Efficient Video Action Recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.096816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.096816Z digest=sha256:02bec571ddb3a65ae61a03b9da971a59a6076483188bfba55769bf728192490d

Observation b1b5eee8-d624-4a21-8933-82b1f20b639e · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vision Transformer Adapter for Dense Predictions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.100923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.100923Z digest=sha256:f3ce7509ae61b3a5b3812cd64376a66ae7d98a4396f7fe50318395e8dd1b99a4

Observation f0189b50-bdb5-4c9c-b43f-64fff44af3f5 · outbound

This paper cites Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Parameter-Efficient Fine-Tuning for Medical Image Analysis: The Missed Opportunity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.105209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.105209Z digest=sha256:773b1b52497ef4a9b4d068142e54d535e892e2db4f4d4140c2fbdcc2a8cdb7e7

Observation 3cf57145-50b6-4376-b42d-feb2aa78d11f · outbound

This paper cites MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.109848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.109848Z digest=sha256:7d3993c7a79a781c9c005c32216f0c82b4720d482b59c1598bd14d3739ad7346

Observation 940aa1de-6d6a-46ff-a5c0-aaa6968df2bc · outbound

This paper cites Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.115036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.115036Z digest=sha256:74b9ee15e46365de005016882b4c7acf4f2e78d58f8672270e920265026732fa

Observation 9c12a0b7-ad35-4518-b3c9-b359c796cb09 · outbound

This paper cites Prompt tuning for parameter- efficient medical image segmentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Prompt tuning for parameter- efficient medical image segmentation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.542188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.119838Z digest=sha256:bf840a628266f51ef792034cacdaeb378346a05adeabcc462203e17bf4c3a144

Observation 12073882-8e25-47fa-b8f3-13c49385940a · outbound

This paper cites Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Towards foundation models and few-shot parameter-efficient fine-tuning for volumetric organ seg- mentation,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.525535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.123971Z digest=sha256:fe853e74ed91744b3e77eef2683ac82a4fd84dadbe63a148631ed15d22023e7d

Observation fd6089af-958e-43a2-98df-a43146bd29da · outbound

This paper cites Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Contrastive Alignment of Vision to Language Through Parameter-Efficient Transfer Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.128128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.128128Z digest=sha256:1f9294a9f834ddb0af481b5f2f1ec18f1f1ec17185f61b19c3ce7e8a1ae4c48c

Observation dfda98a6-9006-4d5c-a902-26e0b1555765 · outbound

This paper cites Scaling language- image pre-training via masking,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Scaling language- image pre-training via masking,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.509160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.132633Z digest=sha256:b907dd48bd511a2e61f25ee011b328c3ef2b8ba46847e6cefdea35e561148ca8

Observation 677245e6-b7da-4913-b2a4-b271c50526f2 · outbound

This paper cites Attention is all you need,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Attention is all you need,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.136469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.136469Z digest=sha256:f28f1c746a48a32c88d2296fb03bf8df9498ed58a7ad533c21a77f89cc38e284

Observation e36c3afa-c4bb-4bf9-af8b-6324a2c96d89 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Gaussian Error Linear Units (GELUs)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.140339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.140339Z digest=sha256:15e7f59223171a5a6aab8c4fdfbd591512d8019940266f955144d1f669b8ffd7

Observation 56ec48d9-6ab1-429d-8320-7828089505f6 · outbound

This paper cites Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Mvco- dot: Multi-view contrastive domain transfer network for medical report generation,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.144569Z digest=sha256:b98c963546b016872c91246f9074a4e3530af8089872a84ba32ebf4628c0b9e7

Observation 88e534ca-45b7-4aef-b061-c2390706973e · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Representation Learning with Contrastive Predictive Coding

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.149064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.149064Z digest=sha256:6388b0056201ba31255cf37d5b12ab163723b69f72555ee75597f93e35b3cc18

Observation ed74d849-4d17-493c-90d0-3a5ca81c09aa · outbound

This paper cites Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:02:21.470237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:02:21.153308Z digest=sha256:cc470aab3574cf49d968591b7bcd5ee973078ce6ef0878def1517e7d0fa94492

Observation c3d7a77f-16a5-4936-be96-df9a5d821719 · outbound

This paper cites Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.157303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.157303Z digest=sha256:050b31f32ca5ec713538eb6d0a0b9b5ca21fd8abbcd85fafbcd4c0d4897e5d10

Observation 01006229-228d-4ad5-9916-e3da28397fdb · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Pytorch: An imperative style, high-performance deep learning library,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.161272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.161272Z digest=sha256:a38ff3ea85b6947c1822da45bcf2c05c74ccd38b64e5c10f1f27069d4b1e55c6

Observation 75493d61-a82c-408a-8cd7-dfe7fe39055b · outbound

This paper cites Decoupled Weight Decay Regularization.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Decoupled Weight Decay Regularization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.165503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.165503Z digest=sha256:1d80d9c238deefc24b2904c1c86e0e8985e2273378e823bc1c2b732191c3e255

Pith citing papers

No inbound Pith citation observations are available.